← Search

Doudou Zhou

3 accepted papers

2026

A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth

ICML 2026poster

Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm. A critical but under-modeled issue is that judge LLMs differ substantially in reliability; treating all judges equally can yield biased leaderboards and misle…

Cited by 0SourceScholar
2026

Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions

ICLR 2026poster

Generalized linear bandits have been extensively studied due to their broad applicability in real-world online decision-making problems. However, these methods typically assume that the expected reward function is known to the users, an assumption that is often unrealistic in practice. Misspecificat…

Cited by 0SourceScholar