2026
CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction
ICML 2026poster
Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's {\em behavioral robustness} to irrelevant or misleading information. In this paper, we argue that a model's true confidence should reflect its stability under cognitive pre…