2026
Sample Smart, Not Hard: Correctness-First Decoding for Better Reasoning in LLMs
ICLR 2026poster
Large Language Models (LLMs) are increasingly applied to complex tasks that require extended reasoning. In such settings, models often benefit from diverse chains-of-thought to arrive at multiple candidate solutions. This requires two competing objectives: to inject enough stochasticity to explore m…