2026
Robust Adaptive Multi-Step Predictive Shielding
ICLR 2026poster
Reinforcement learning for safety-critical tasks requires policies that are both high-performing and safe throughout the learning process. While model-predictive shielding is a promising approach, existing methods are often computationally intractable for the high-dimensional, nonlinear systems wher…