2025
CourtEval: A Courtroom-Based Multi-Agent Evaluation Framework
ACL 2025finding
Automated evaluation is crucial for assessing the quality of natural language text, especially in open-ended generation tasks, given the costly and time-consuming nature of human evaluation. Existing automatic evaluation metrics like ROUGE and BLEU often show low correlation with human judgments. As…