← Search

Xinxuan Lu

1 accepted papers

2026

Camera Control for Text-to-Image Generation via Learning Viewpoint Tokens

CVPR 2026

Current text-to-image models struggle to provide precise camera control using natural language alone. In this work, we present a framework for precise camera control with global scene understanding in text-to-image generation by learning parametric camera tokens. We fine-tune image generation models

Cited by 0SourcecodeScholar