2025
SynthesizeMe! Inducing Persona-Guided Prompts for Personalized Reward Models in LLMs
ACL 2025long
Recent calls for pluralistic alignment of Large Language Models (LLMs) encourage adapting models to diverse user preferences. However, most prior work on personalized reward models heavily rely on additional identity information, such as demographic details or a predefined set of preference categori…