2024
A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback
ICML 2024poster
Inverse Reinforcement Learning (IRL) and Reinforcement Learning from Human Feedback (RLHF) are pivotal methodologies in reward learning, which involve inferring and shaping the underlying reward function of sequential decision-making problems based on observed human demonstrations and feedback. Most…