2026
Offline Multi-Agent Reinforcement Learning via Sequential Score Decomposition
ICML 2026poster
Offline cooperative multi-agent reinforcement learning (MARL) faces unique challenges due to the distribution shift between online and offline data collection. While online MARL typically converges to a single coordinated joint policy, offline datasets are often mixtures of diverse cooperative behav…