← Search

Tiantian Zhang

9 accepted papers

2026

Principled RL for Flow Matching Emerges From the Chunk-level Policy Optimization

ICML 2026poster

Recent Progress in post-training flow matching for text-to-image (T2I) generation with Group Relative Policy Optimization (GRPO) has demonstrated strong potential. However, it is hindered by a critical limitation: inaccurate advantage attribution. In this work, we argue that aggregating consecutive …

Cited by 0SourceScholar
2025

Hybrid Reciprocal Transformer with Triplet Feature Alignment for Scene Graph Generation

CVPR 2025poster

Scene graph generation is a pivotal task in computer vision, focusing on comprehensive identification of visual relation tuples embedded within images. The advancement of methods involving triplets has sought to enhance task performance by integrating triplets as contextual features for more precise…

2025

Monocular Visual-Inertial Odometry Based on Local Maximum A Posteriori Estimation

RA-L 2025

Monocular visual-inertial odometry can effectively solve the problem of unobservable metric scale in monocular odometry. However, due to differences in sensor observation noise and redundant constraints, the positioning accuracy of a general monocular visual-inertial odometer may decline compared wi

Cited by 2SourceScholar
2025

Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation

NeurIPS 2025poster

Reinforcement learning (RL) has garnered increasing attention in text-to-image (T2I) generation. However, most existing RL approaches are tailored to either diffusion models or autoregressive models, overlooking an important alternative: masked generative models. In this work, we propose Mask-GRPO,…

Cited by 0SourceScholar
2023

CCVO: Cascaded CNNs for Fast Monocular Visual Odometry Towards the Dynamic Environment

RA-L 2023

For AR applications, the present monocular VO methods can be further improved to provide more real-time and accurate self-localization in a dynamic environment with motion disturbance. This letter proposes the CCVO (Cascaded CNNs for Visual Odometry) which is a monocular VO approach to realize end-t

Cited by 10SourceScholar
2023

Curriculum-based Co-design of Morphology and Control of Voxel-based Soft Robots

ICLR 2023poster

Co-design of morphology and control of a Voxel-based Soft Robot (VSR) is challenging due to the notorious bi-level optimization. In this paper, we present a Curriculum-based Co-design (CuCo) method for learning to design and control VSRs through an easy-to-difficult process. Specifically, we expand…

Cited by 10SourcePDFScholar
2023

PreCo: Enhancing Generalization in Co-Design of Modular Soft Robots via Brain-Body Pre-Training

CoRL 2023oral

Brain-body co-design, which involves the collaborative design of control strategies and morphologies, has emerged as a promising approach to enhance a robot's adaptability to its environment. However, the conventional co-design process often starts from scratch, lacking the utilization of prior know…

Cited by 9SourceScholar