2025
Probabilistic Shielding for Safe Reinforcement Learning
AAAI 2025technical
In real-life scenarios, a Reinforcement Learning (RL) agent aiming to maximize their reward, must often also behave in a safe manner, including at training time. Thus, much attention in recent years has been given to Safe RL, where an agent aims to learn an optimal policy among all policies that sat…