← Search

Xiangyu Bai

4 accepted papers

2026

MoReGen: Multi-Agent Motion-Reasoning Engine for Code-based Text-to-Video Synthesis

CVPR 2026

While text-to-video (T2V) generation has achieved remarkable progress in photorealism, generating intent-aligned videos that faithfully obey physics principles remains a core challenge. In this work, we systematically study Newtonian motion-controlled text-to-video generation and evaluation, emphasi

Cited by 0SourcecodeScholar
2026

UniTrack: Differentiable Graph Representation Learning for Multi-Object Tracking

ICLR 2026poster

We present UniTrack, a plug-and-play graph-theoretic loss function designed to significantly enhance multi-object tracking (MOT) performance by directly optimizing tracking-specific objectives through unified differentiable learning. Unlike prior graph-based MOT methods that redesign tracking archit…

Cited by 0SourcecodeScholar
2023

An Evaluation Platform to Scope Performance of Synthetic Environments in Autonomous Ground Vehicles Simulation

ICASSP 2023accepted

Evaluating autonomous ground vehicles requires evaluating their mobility performance. Since autonomous vehicles are envisioned to make decisions in a variety of situations and environments too diverse to practically assess with only physical testing, their development, and evaluation will necessaril…

Cited by 0SourceScholar