2026
Multi-Value Alignment for LLMs via Value Decorrelation and Extrapolation
AAAI 2026technical
With the rapid advancement of large language models (LLMs), aligning them with human values for safety and ethics has become a critical challenge. This problem is especially challenging when multiple, potentially conflicting human values must be considered and balanced. Although several variants of