2026
Constrained Meta Reinforcement Learning with Provable Test-Time Safety
ICML 2026poster
Meta reinforcement learning (RL) allows agents to leverage experience across a distribution of tasks on which the agent can *train* at will, enabling faster learning of optimal policies on new *test* tasks. Despite its success in improving sample complexity on test tasks, many real-world application…