← Search

Parsa Mirtaheri

2 accepted papers

2025

Direct Alignment with Heterogeneous Preferences

NeurIPS 2025poster

Alignment with human preferences is commonly framed using a universal reward function, even though human preferences are inherently heterogeneous. We formalize this heterogeneity by introducing user types and examine the limits of the homogeneity assumption. We show that aligning to heterogeneous pr…

Cited by 0SourcecodeScholar
2025

Let Me Think! A Long Chain of Thought Can Be Worth Exponentially Many Short Ones

NeurIPS 2025poster

Inference-time computation has emerged as a promising scaling axis for improving large language model reasoning. However, despite yielding impressive performance, the optimal allocation of inference-time computation remains poorly understood. A central question is whether to prioritize sequential sc…

Cited by 0SourcecodeScholar