2024
LLM Evaluators Recognize and Favor Their Own Generations
NeurIPS 2024oral
Self-evaluation using large language models (LLMs) has proven valuable not only in benchmarking but also methods like reward modeling, constitutional AI, and self-refinement. But new biases are introduced due to the same LLM acting as both the evaluator and the evaluatee. One such bias is self-prefe…