2026
Prototype Entropy Alignment: Reinforcing Structured Uncertainty in LLM Reasoning
AAAI 2026technical
Recent research reveals that a minority of high-entropy tokens significantly influence the reasoning quality of large language models (LLMs). Inspired by this, we propose Prototype Entropy Alignment (PEA), a reinforcement learning framework that models effective reasoning not as a single path but as