2025
A Ranking Scheme for Trust Region Multi-agent Reinforcement Learning
ICASSP 2025accepted
In multi-agent reinforcement learning (MARL), trust region (TR) methods are widely used because they effectively mitigate the nonstationarity of multi-agent systems and facilitate collaboration among diverse agent types. Based on the multi-agent advantage decomposition lemma, TR methods adopt a sequ…