2025
Benchmarking and Improving LLM Robustness for Personalized Generation
EMNLP 2025
Recent years have witnessed a growing interest in personalizing the responses of large language models (LLMs). While existing evaluations primarily focus on whether a response aligns with a user’s preferences, we argue that factuality is an equally important yet often overlooked dimension. In the co