2025
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
EMNLP 2025
LLM-as-Judge frameworks are increasingly popular for AI evaluation, yet research findings on the relationship between models’ generation and judgment abilities remain inconsistent. We investigate this relationship through systematic dataset- and instance-level analyses across 11 models and 21 divers