2021
QPLEX: Duplex Dueling Multi-Agent Q-Learning
ICLR 2021poster
We explore value-based multi-agent reinforcement learning (MARL) in the popular paradigm of centralized training with decentralized execution (CTDE). CTDE has an important concept, Individual-Global-Max (IGM) principle, which requires the consistency between joint and local action selections to supp…