← Search

Qisong Yang

3 accepted papers

2024

Analyzing Generalization in Policy Networks: A Case Study with the Double-Integrator System

AAAI 2024technical

Extensive utilization of deep reinforcement learning (DRL) policy networks in diverse continuous control tasks has raised questions regarding performance degradation in expansive state spaces where the input state norm is larger than that in the training environment. This paper aims to uncover the u…

2021

WCSAC: Worst-Case Soft Actor Critic for Safety-Constrained Reinforcement Learning

AAAI 2021technical

Safe exploration is regarded as a key priority area for reinforcement learning research. With separate reward and safety signals, it is natural to cast it as constrained reinforcement learning, where expected long-term costs of policies are constrained. However, it can be hazardous to set constraint…