2024
ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate
ICLR 2024poster
Text evaluation has historically posed significant challenges, often demanding substantial labor and time cost. With the emergence of large language models (LLMs), researchers have explored LLMs' potential as alternatives for human evaluation. While these single-agent-based approaches show promise,…