2022
Robust Imitation via Mirror Descent Inverse Reinforcement Learning
NeurIPS 2022accept
Recently, adversarial imitation learning has shown a scalable reward acquisition method for inverse reinforcement learning (IRL) problems. However, estimated reward signals often become uncertain and fail to train a reliable statistical model since the existing methods tend to solve hard optimizatio…