2025
Understanding How Value Neurons Shape the Generation of Specified Values in LLMs
EMNLP 2025
Rapid integration of large language models (LLMs) into societal applications has intensified concerns about their alignment with universal ethical principles, as their internal value representations remain opaque despite behavioral alignment advancements. Current approaches struggle to systematicall