2023
A Perturbation-Based Policy Distillation Framework with Generative Adversarial Nets
ICASSP 2023accepted
We study the problem of imitation learning in automated decision systems, in which a learner is trained to imitate an expert demonstrator. A widely used method is adversarial imitation learning that alternately optimizes a generator (learner) and a discriminator (reward function). However, the discr…