2021
Uncertainty-Aware Policy Optimization: A Robust, Adaptive Trust Region Approach
AAAI 2021technical
In order for reinforcement learning techniques to be useful in real-world decision making processes, they must be able to produce robust performance from limited data. Deep policy optimization methods have achieved impressive results on complex tasks, but their real-world adoption remains limited be…