2026
Geometric Control of Out-of-Distribution Shift in Safe Offline RL
ICML 2026poster
Safe offline reinforcement learning (RL) requires optimizing policies within the support of static datasets while satisfying strict safety constraints. Although recent latent generative policies achieve strong empirical performance, they rely heavily on implicit regularization and lack systematic co…