2025
AHa-Bench: Benchmarking Audio Hallucinations in Large Audio-Language Models
NeurIPS 2025poster
Hallucinations present a significant challenge in the development and evaluation of large language models (LLMs), directly affecting their reliability and accuracy. While notable advancements have been made in research on textual and visual hallucinations, there is still a lack of a comprehensive be…