2026
Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning
CVPR 2026
Video reasoning has advanced with large multimodal models (LMMs), yet their inference is often a single pass that returns an answer without verifying whether the reasoning is evidence-aligned. We introduce **Reinforce to Learn, Elect to Reason (RLER)**, a dual paradigm that decouples learning to pro