2025
VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
ICCV 2025poster
Video generation models have achieved remarkable progress in text-to-video tasks. These models are typically trained on text-video pairs with highly detailed and carefully crafted descriptions, while real-world user inputs during inference are often concise, vague, or poorly structured. This gap mak…