2026
Video-KTR: Reinforcing Video Reasoning via Key Token Attribution
ICLR 2026poster
Reinforcement learning (RL) has shown strong potential for enhancing reasoning in multimodal large language models (MLLMs), yet existing video reasoning methods often rely on coarse sequence-level rewards or single-factor token selection. Such approaches neglect fine-grained links among visual input…