2026
Do LLMs Signal When They’re Right? Evidence from Neuron Agreement
ICML 2026spotlight
Large language models (LLMs) commonly boost reasoning via sample-evaluate-ensemble decoders (e.g., majority voting), achieving label free gains without ground truth. However, prevailing strategies score candidates using only external outputs such as token probabilities, entropies, or self evaluation…