← Search

Mohammad Anas Jawad

1 accepted papers

2026

CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction

ICML 2026poster

Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's {\em behavioral robustness} to irrelevant or misleading information. In this paper, we argue that a model's true confidence should reflect its stability under cognitive pre…

Cited by 0SourceScholar