← Search

Prajwal Koirala

3 accepted papers

2025

Latent Safety-Constrained Policy Approach for Safe Offline Reinforcement Learning

ICLR 2025poster

In safe offline reinforcement learning, the objective is to develop a policy that maximizes cumulative rewards while strictly adhering to safety constraints, utilizing only offline data. Traditional methods often face difficulties in balancing these constraints, leading to either diminished performa…