2025
Learning on One Mode: Addressing Multi-modality in Offline Reinforcement Learning
ICLR 2025poster
Offline reinforcement learning (RL) seeks to learn optimal policies from static datasets without interacting with the environment. A common challenge is handling multi-modal action distributions, where multiple behaviours are represented in the data. Existing methods often assume unimodal behaviour…