← Search

Pierluigi Vito Amadori

1 accepted papers

2019

Random Expert Distillation: Imitation Learning via Expert Policy Support Estimation

ICML 2019oral

We consider the problem of imitation learning from a finite set of expert trajectories, without access to reinforcement signals. The classical approach of extracting the expert’s reward function via inverse reinforcement learning, followed by reinforcement learning is indirect and may be computation…