AIR: Unifying Individual and Collective Exploration in Cooperative Multi-Agent Reinforcement Learning
Exploration in cooperative multi-agent reinforcement learning (MARL) remains challenging for value-based agents due to the absence of an explicit policy. Existing approaches include individual exploration based on uncertainty towards the system and collective exploration through behavioral diversity…