2020
Hybrid Learning for Multi-agent Cooperation with Sub-optimal Demonstrations
IJCAI 2020poster
This paper aims to learn multi-agent cooperation where each agent performs its actions in a decentralized way. In this case, it is very challenging to learn decentralized policies when the rewards are global and sparse. Recently, learning from demonstrations (LfD) provides a promising way to handle…