← Search

Luise Ge

4 accepted papers

2025

Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks

ICML 2025poster

Many dynamic decision problems, such as robotic control, involve a series of tasks, many of which are unknown at training time. Typical approaches for these problems, such as multi-task and meta reinforcement learning, do not generalize well when the tasks are diverse. On the other hand, approaches…

2024

Axioms for AI Alignment from Human Feedback

NeurIPS 2024spotlight

In the context of reinforcement learning from human feedback (RLHF), the reward function is generally derived from maximum likelihood estimation of a random utility model based on pairwise comparisons made by humans. The problem of learning a reward function is one of preference aggregation that, we…

Cited by 17SourcePDFScholar