2025
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
NeurIPS 2025spotlight
Despite the remarkable reasoning performance, eliciting the long chain-of-thought(CoT) ability in large language models(LLMs) typically requires costly reinforcement learning or supervised fine-tuning on high-quality distilled data. We investigate the internal mechanisms behind this capability and s…