← Search

Jongwon Lim

1 accepted papers

2026

Dual Mechanisms of Value Expression: Intrinsic vs. Prompted Values in Large Language Models

ICML 2026poster

Large language models can express values in two main ways: (1) $\textit{intrinsic}$ expression, reflecting the model's inherent values learned during training, and (2) $\textit{prompted}$ expression, elicited by explicit prompts. Given their widespread use in value alignment, it is paramount to clea…

Cited by 0SourceScholar