2026
Offline Reinforcement Learning with Adaptive Feature Fusion
ICLR 2026poster
Return-conditioned supervised learning (RCSL) algorithms have demonstrated strong generative capabilities in offline reinforcement learning (RL) by learning action distributions based on both the state and the return. However, many existing approaches treat RL as a conditional sequence modeling task…