2026
Mind the Gap: Catching Hallucinations via Evidence Drop on the Reasoning Manifold
ICML 2026poster
Large Language Models (LLMs) show strong reasoning abilities, yet their reliability is hindered by hallucinations, where fluent reasoning becomes factually or logically incorrect. Most existing uncertainty-based detectors rely on sequence-level averaging, which ignores the step-wise dynamics of reas…