2026
The Geometry of Reasoning: Self-Evaluation via Layerwise Trajectory Evolution
ICML 2026poster
Large Reasoning Models (LRMs) enhance performance by generating explicit Chain-of-Thought (CoT) trajectories, yet enabling them to self-evaluate correctness without external supervision remains a critical challenge. Existing methods often rely on ground-truth labels or shallow output probabilities, …