← Search

Silang Wu

1 accepted papers

2025

FOSP: Fine-tuning Offline Safe Policy through World Models

ICLR 2025poster

Offline Safe Reinforcement Learning (RL) seeks to address safety constraints by learning from static datasets and restricting exploration. However, these approaches heavily rely on the dataset and struggle to generalize to unseen scenarios safely. In this paper, we aim to improve safety during the d…