← Search

Mingxi Zou

3 accepted papers

2026

DARC: Disagreement-Aware Alignment via Risk-Constrained Decoding

ICML 2026poster

Preference-based alignment methods (e.g., RLHF, DPO) typically optimize a single scalar objective, implicitly averaging over heterogeneous human preferences. In practice, systematic annotator and user-group disagreement makes mean-reward maximization brittle and susceptible to proxy over-optimizatio…

Cited by 0SourceScholar
2025

FinHEAR: Human Expertise and Adaptive Risk-Aware Temporal Reasoning for Financial Decision-Making

EMNLP 2025

Financial decision-making presents unique challenges for language models, requiring them to handle temporally evolving, risk-sensitive, and event-driven contexts. While large language models (LLMs) demonstrate strong general reasoning abilities, they often overlook key behavioral patterns underlying

2025

From Implicit Exploration to Structured Reasoning: Guideline and Refinement for LLMs

EMNLP 2025

Large language models (LLMs) have advanced general-purpose reasoning, showing strong performance across diverse tasks. However, existing methods often rely on implicit exploration, where the model follows stochastic and unguided reasoning paths—like walking without a map. This leads to unstable reas

Cited by 0SourcePDFScholar