2025
Robust Hallucination Detection in LLMs via Adaptive Token Selection
NeurIPS 2025poster
Hallucinations in large language models (LLMs) pose significant safety concerns that impede their broader deployment. Recent research in hallucination detection has demonstrated that LLMs' internal representations contain truthfulness hints, which can be harnessed for detector training. However, the…