← Search

Arsh Zahed

1 accepted papers

2019

On-Policy Robot Imitation Learning from a Converging Supervisor

CoRL 2019

Existing on-policy imitation learning algorithms, such as DAgger, assume access to a fixed supervisor. However, there are many settings where the supervisor may evolve during policy learning, such as a human performing a novel task or an improving algorithmic controller. We formalize imitation learn