← Search

Mehdi Mirzazadeh

1 accepted papers

2026

Distribution-Calibrated Inference Time Compute for Thinking LLM-as-a-Judge

ICML 2026poster

Thinking Large Language Models (LLMs) used as judges for pairwise preferences remain noisy at the single-sample level, and common aggregation rules (majority vote, soft self-consistency, or instruction-based self-aggregation) are inconsistent when ties are allowed. We study inference-time compute (I…

Cited by 0SourceScholar