← Search

Peter F. Ahnn

1 accepted papers

2026

Learning to summarize user information for personalized reinforcement learning from human feedback

ICLR 2026poster

As everyday use cases of large language model (LLM) AI assistants have expanded, it is becoming increasingly important to personalize responses to align to different users' preferences and goals. While reinforcement learning from human feedback (RLHF) is effective at improving LLMs to be generally m…

Cited by 0SourceScholar