CoRL 2024poster3 citations

Exploring Under Constraints with Model-Based Actor-Critic and Safety Filters

Ahmed Agha, Baris Kayalibay, Atanas Mirchev, Patrick van der Smagt, Justin Bayer

Abstract

Applying reinforcement learning (RL) to learn effective policies on physical robots without supervision remains challenging when it comes to tasks where safe exploration is critical. Constrained model-based RL (CMBRL) presents a promising approach to this problem. These methods are designed to learn constraint-adhering policies through constrained optimization approaches. Yet, such policies often fail to meet stringent safety requirements during learning and exploration. Our solution ``CASE'' aims to reduce the instances where constraints are breached during the learning phase. Specifically, CASE integrates techniques for optimizing constrained policies and employs planning-based safety filters as backup policies, effectively lowering constraint violations during learning and making it a more reliable option than other recent constrained model-based policy optimization methods.

Model-based RLSafe RLSafety FilterExploration
BibTeX
@inproceedings{
agha2024exploring,
title={Exploring Under Constraints with Model-Based Actor-Critic and Safety Filters},
author={Ahmed Agha and Baris Kayalibay and Atanas Mirchev and Patrick van der Smagt and Justin Bayer},
booktitle={8th Annual Conference on Robot Learning},
year={2024},
url={https://openreview.net/forum?id=s31IWg2kN5}
}
Exploring Under Constraints with Model-Based Actor-Critic and Safety Filters · CoRL 2024