2024
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
NeurIPS 2024poster
Among approaches for provably safe reinforcement learning, Model Predictive Shielding (MPS) has proven effective at complex tasks in continuous, high-dimensional state spaces, by leveraging a *backup policy* to ensure safety when the learned policy attempts to take risky actions. However, while MPS…