← Search

Pablo Parrilo

2 accepted papers

2024

A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback

ICML 2024poster

Inverse Reinforcement Learning (IRL) and Reinforcement Learning from Human Feedback (RLHF) are pivotal methodologies in reward learning, which involve inferring and shaping the underlying reward function of sequential decision-making problems based on observed human demonstrations and feedback. Most…

Cited by 1SourcePDFScholar
2024

Towards Tight Convex Relaxations for Contact-Rich Manipulation

RSS 2024poster

We present a novel method for global motion planning of robotic systems that interact with the environment through contacts. Our method directly handles the hybrid nature of such tasks using tools from convex optimization. We formulate the motion-planning problem as a shortest-path problem in a grap…