2026
VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
ICML 2026oral
Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to capture deeper motivational principles. Value-based approaches offer a more principled path, yet three gaps persist– extraction often ignores hierarchic…