2025
Beyond the Surface: Enhancing LLM-as-a-Judge Alignment with Human via Internal Representations
NeurIPS 2025poster
The growing scale of evaluation tasks has led to the widespread adoption of automated evaluation using LLMs, a paradigm known as “LLM-as-a-judge”. However, improving its alignment with human preferences without complex prompts or fine-tuning remains challenging. Previous studies mainly optimize base…