2024
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained Generation
ACL 2024long
Large language models (LLMs) produce hallucinated text, compromising their practical utility in professional contexts. To assess the reliability of LLMs, numerous initiatives have developed benchmark evaluations for hallucination phenomena. However, they often employ constrained generation technique…