2026
FairJudge : An Adaptive, Debiased, and Consistent LLM-as-a-Judge
ICML 2026poster
Existing LLM-as-a-Judge systems suffer from three fundamental limitations: \textbf{limited adaptivity} to task and domain-specific evaluation criteria, \textbf{systematic biases} driven by non-semantic cues such as position, length, format, and model provenance, and \textbf{evaluation inconsistency}…