2025
Influences on LLM Calibration: A Study of Response Agreement, Loss Functions, and Prompt Styles
ACL 2025long
Calibration, the alignment between model confidence and prediction accuracy, is critical for the reliable deployment of large language models (LLMs). Existing works neglect to measure the generalization of their methods to other prompt styles and different sizes of LLMs. To address this, we define a…