Learning Noise-Induced Reward Functions for Surpassing Demonstrations in Imitation Learning
Imitation learning (IL) has recently shown impressive performance in training a reinforcement learning agent with human demonstrations, eliminating the difficulty of designing elaborate reward functions in complex environments. However, most IL methods work under the assumption of the optimality of…