2026
Gradient-Protected Value Decomposition for Cooperative Multi-Agent Reinforcement Learning
AAAI 2026technical
In recent years, deep multi-agent reinforcement learning (MARL) has demonstrated remarkable potential in solving complex cooperative tasks by enabling decentralized yet efficient coordination among agents. However, during decentralized training, agent policy updates induced by different joint action