← Search

Zezhong Qian

2 accepted papers

2026

Predicting What Matters: Robust Generalist Robot Policy Learning via Future Semantic Mask

ICML 2026poster

World models derived from large-scale video generative pre-training have emerged as a promising paradigm for generalist robot policy learning. However, standard approaches often focus on high-fidelity RGB video prediction, but this can result in overfitting to irrelevant factors, such as dynamic bac…

Cited by 0SourceScholar
2025

Dualdiff: Dual-Branch Diffusion Model for Autonomous Driving with Semantic Fusion

ICRA 2025

Accurate and high-fidelity driving scene reconstruction relies on fully leveraging scene information as conditioning. However, existing approaches, which primarily use 3D bounding boxes and binary maps for foreground and background control, fall short in capturing the complexity of the scene and int

Cited by 5SourceScholar