← Search

Guiyu Zhang

3 accepted papers

2026

LIVE: Long-horizon Interactive Video World Modeling

ICML 2026poster

Autoregressive video world models predict future visual observations conditioned on actions. While effective over short horizons, these models often struggle with long-horizon generation, as small prediction errors accumulate over time. Prior methods alleviate this by introducing pre-trained teacher…

Cited by 0SourceScholar
2026

SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation

CVPR 2026

Controlling both camera motion and object dynamics is essential for coherent and expressive video generation, yet current methods typically handle only one motion type or rely on ambiguous 2D cues that entangle camera-induced parallax with true object movement. We present SymphoMotion, a unified mot

Cited by 1SourceScholar
2025

Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward Modeling

ICLR 2025poster

In this paper, we focus on the task of conditional image generation, where an image is synthesized according to user instructions. The critical challenge underpinning this task is ensuring both the fidelity of the generated images and their semantic alignment with the provided conditions. To tackle…

Cited by 4SourcePDFScholar