← Search

Xun Shen

5 accepted papers

2025

Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies

NeurIPS 2025spotlight

When applying offline reinforcement learning (RL) in healthcare scenarios, the out-of-distribution (OOD) issues pose significant risks, as inappropriate generalization beyond clinical expertise can result in potentially harmful recommendations. While existing methods like conservative Q-learning (C…

Cited by 0SourceScholar
2024

A Survey of Constraint Formulations in Safe Reinforcement Learning

IJCAI 2024poster

Safety is critical when applying reinforcement learning (RL) to real-world problems. As a result, safe RL has emerged as a fundamental and powerful paradigm for optimizing an agent’s policy while incorporating notions of safety. A prevalent safe RL approach is based on a constrained criterion, which…

2024

Flipping-based Policy for Chance-Constrained Markov Decision Processes

NeurIPS 2024poster

Safe reinforcement learning (RL) is a promising approach for many real-world decision-making problems where ensuring safety is a critical necessity. In safe RL research, while expected cumulative safety constraints (ECSCs) are typically the first choices, chance constraints are often more pragmatic…

Cited by 1SourcePDFScholar
2023

Safe Exploration in Reinforcement Learning: A Generalized Formulation and Algorithms

NeurIPS 2023poster

Safe exploration is essential for the practical use of reinforcement learning (RL) in many real-world scenarios. In this paper, we present a generalized safe exploration (GSE) problem as a unified formulation of common safe exploration problems. We then propose a solution of the GSE problem in the f…

Cited by 13SourcePDFScholar
2020

Cooperative Comfortable-Driving at Signalized Intersections for Connected and Automated Vehicles

RA-L 2020

This letter proposes a control framework for Connected and Automated Vehicles(CAVs) to approach the signalized intersections with good driving-comfortability. Both the velocity plan and longitudinal dynamics control are concerned in this study. Regarding the velocity plan problem, a two-layer framew

Cited by 36SourceScholar