2025
LLMs are Biased Evaluators But Not Biased for Fact-Centric Retrieval Augmented Generation
ACL 2025finding
Recent studies have demonstrated that large language models (LLMs) exhibit significant biases in evaluation tasks, particularly in preferentially rating and favoring self-generated content. However, the extent to which this bias manifests in fact-oriented tasks, especially within retrieval-augmented…