← Search

Rundong Shi

1 accepted papers

2024

Calibrating the Confidence of Large Language Models by Eliciting Fidelity

EMNLP 2024main

Large language models optimized with techniques like RLHF have achieved good alignment in being helpful and harmless. However, post-alignment, these language models often exhibit overconfidence, where the expressed confidence does not accurately calibrate with their correctness rate. In this paper,…

Cited by 4SourcePDFScholar