← Search

Chufan Chen

1 accepted papers

2025

Rebalancing Return Coverage for Conditional Sequence Modeling in Offline Reinforcement Learning

NeurIPS 2025poster

Recent advancements in offline reinforcement learning (RL) have underscored the capabilities of conditional sequence modeling (CSM), a paradigm that models the action distribution conditioned on both historical trajectories and target returns associated with each state. However, due to the imbalance…

Cited by 0SourceScholar