2026
SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning
CVPR 2026
As humans, we are natural any-horizon reasoners, i.e., we can decide whether to iteratively skim long videos or watch short ones in full when necessary for a given task. With this in mind, one would expect video reasoning models to reason flexibly across different durations. However, SOTA models are