2022
Learning Cooperative Multi-Agent Policies With Partial Reward Decoupling
RA-L 2022
One of the preeminent obstacles to scaling multi-agent reinforcement learning to large numbers of agents is assigning credit to individual agents’ actions. In this letter, we address this credit assignment problem with an approach that we call <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" x