2024
Calibrating the Confidence of Large Language Models by Eliciting Fidelity
EMNLP 2024main
Large language models optimized with techniques like RLHF have achieved good alignment in being helpful and harmless. However, post-alignment, these language models often exhibit overconfidence, where the expressed confidence does not accurately calibrate with their correctness rate. In this paper,…