2021
Robust Reinforcement Learning using Least Squares Policy Iteration with Provable Performance Guarantees
ICML 2021spotlight
This paper addresses the problem of model-free reinforcement learning for Robust Markov Decision Process (RMDP) with large state spaces. The goal of the RMDPs framework is to find a policy that is robust against the parameter uncertainties due to the mismatch between the simulator model and real-wor…