2026
On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models
ICML 2026spotlight
Recent reinforcement learning (RL) techniques have yielded impressive reasoning improvements in language models, yet it remains unclear whether post-training truly extends a model’s reasoning ability beyond what it acquires during pre-training. A central challenge is the lack of control in modern tr…