2025
LongVPO: From Anchored Cues to Self-Reasoning for Long-Form Video Preference Optimization
NeurIPS 2025poster
We present LongVPO, a novel two‑stage Direct Preference Optimization framework that enables short‑context vision‑language models to robustly understand ultra‑long videos without any long‑video annotations. In Stage 1, we synthesize preference triples by anchoring questions to individual short clips,…