2026
More Thought, Less Accuracy? On the Dual Nature of Reasoning in Vision-Language Models
ICLR 2026poster
Reasoning has emerged as a pivotal capability in Large Language Models (LLMs). Through Reinforcement Learning (RL), typically Group Relative Policy Optimization (GRPO), these models are able to solve complex tasks such as mathematics and code generation. Building on these advances, recent research h…