← Search

Shanchao Yang

1 accepted papers

2026

Offline Multi-Agent Reinforcement Learning via Sequential Score Decomposition

ICML 2026poster

Offline cooperative multi-agent reinforcement learning (MARL) faces unique challenges due to the distribution shift between online and offline data collection. While online MARL typically converges to a single coordinated joint policy, offline datasets are often mixtures of diverse cooperative behav…

Cited by 0SourceScholar