← Search

Mianchu Wang

2 accepted papers

2025

Learning on One Mode: Addressing Multi-modality in Offline Reinforcement Learning

ICLR 2025poster

Offline reinforcement learning (RL) seeks to learn optimal policies from static datasets without interacting with the environment. A common challenge is handling multi-modal action distributions, where multiple behaviours are represented in the data. Existing methods often assume unimodal behaviour…