2026
The ACE Protocol: Operationalizing Language Model Activations for Better Calibration and Utility
ICML 2026poster
As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good proxy for trust: well-calibrated confidence estimates help inform the risk versus reward trade-off when trusting a specific model output. Unfortunately, e…