← Search

Hamid Dadkhahi

5 accepted papers

2026

Distribution-Calibrated Inference Time Compute for Thinking LLM-as-a-Judge

ICML 2026poster

Thinking Large Language Models (LLMs) used as judges for pairwise preferences remain noisy at the single-sample level, and common aggregation rules (majority vote, soft self-consistency, or instruction-based self-aggregation) are inconsistent when ties are allowed. We study inference-time compute (I…

Cited by 0SourceScholar
2025

Learning from others' mistakes: Finetuning machine translation models with span-level error annotations

ICML 2025poster

Despite growing interest in incorporating feedback to improve language models, most efforts focus only on sequence-level annotations. In this work, we explore the potential of utilizing fine-grained span-level annotations from offline datasets to improve model quality. We develop a simple finetuning…

Cited by 1SourcePDFScholar
2023

Order Matters in the Presence of Dataset Imbalance for Multilingual Learning

NeurIPS 2023poster

In this paper, we empirically study the optimization dynamics of multi-task learning, particularly focusing on those that govern a collection of tasks with significant data imbalance. We present a simple yet effective method of pre-training on high-resource tasks, followed by fine-tuning on a mixtur…

Cited by 7SourcePDFScholar
2022

Fourier Representations for Black-Box Optimization over Categorical Variables

AAAI 2022technical

Optimization of real-world black-box functions defined over purely categorical variables is an active area of research. In particular, optimization and design of biological sequences with specific functional or structural properties have a profound impact in medicine, materials science, and biotechn…

Cited by 9SourcePDFScholar