Transformer-Based Multi-Agent Reinforcement Learning Method With Credit-Oriented Strategy Differentiation
The problem of Multi-Agent Reinforcement Learning (MARL) shows a high level of both complexity in the environment and coordination between agents. In order to scale the algorithm to large-scale agent scenarios, neural networks designed for MARL are typically implemented with parameter sharing. These