2021
Guided Exploration with Proximal Policy Optimization using a Single Demonstration
ICML 2021spotlight
Solving sparse reward tasks through exploration is one of the major challenges in deep reinforcement learning, especially in three-dimensional, partially-observable environments. Critically, the algorithm proposed in this article is capable of using a single human demonstration to solve hard-explora…