← Search

Sebastian Lopez-Cot

1 accepted papers

2021

A Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning

ICML 2021spotlight

A fundamental challenge in multiagent reinforcement learning is to learn beneficial behaviors in a shared environment with other simultaneously learning agents. In particular, each agent perceives the environment as effectively non-stationary due to the changing policies of other agents. Moreover, e…