2024
Aligning LLM Agents by Learning Latent Preference from User Edits
NeurIPS 2024poster
We study interactive learning of language agents based on user edits made to the agent's output. In a typical setting such as writing assistants, the user interacts with a language agent to generate a response given a context, and may optionally edit the agent response to personalize it based on the…