2025
HEAL: A Hypothesis-Based Preference-Aware Analysis Framework
EMNLP 2025
Preference optimization methods like DPO have achieved remarkable performance in LLM alignment. However, the evaluation for these methods relies on a single response and overlooks other potential outputs, which could also be generated in real-world applications within this hypothetical space. To add