2021
Safe Reinforcement Learning Using Advantage-Based Intervention
ICML 2021spotlight
Many sequential decision problems involve finding a policy that maximizes total reward while obeying safety constraints. Although much recent research has focused on the development of safe reinforcement learning (RL) algorithms that produce a safe policy after training, ensuring safety during train…