2026
Mitigating the Safety–Utility Trade-off in LLM Alignment via Adaptive Safe Context Learning
ICML 2026poster
While reasoning models have achieved remarkable success in complex reasoning tasks, their increasing power necessitates stringent safety measures. For safety alignment, the core challenge lies in the inherent trade-off between safety and utility. However, prevailing alignment strategies typically co…