Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
Uncertainty estimation (UE) of generative large language models (LLMs) is crucial for evaluating the reliability of generated sequences. A significant subset of UE methods utilize token probabilities to assess uncertainty, aggregating multiple token probabilities into a single UE score using a scori…