2022
Independent Natural Policy Gradient always converges in Markov Potential Games
AISTATS 2022poster
Natural policy gradient has emerged as one of the most successful algorithms for computing optimal policies in challenging Reinforcement Learning (RL) tasks, yet, very little was known about its convergence properties until recently. The picture becomes more blurry when it comes to multi-agent RL (M…