2025
M-MAD: Multidimensional Multi-Agent Debate for Advanced Machine Translation Evaluation
ACL 2025long
Recent advancements in large language models (LLMs) have given rise to the LLM-as-a-judge paradigm, showcasing their potential to deliver human-like judgments. However, in the field of machine translation (MT) evaluation, current LLM-as-a-judge methods fall short of learned automatic metrics. In thi…