2021
MAMBPO: Sample-efficient multi-robot reinforcement learning using learned world models
IROS 2021poster
Multi-robot systems can benefit from reinforcement learning (RL) algorithms that learn behaviours in a small number of trials, a property known as sample efficiency. This research thus investigates the use of learned world models to improve sample efficiency. We present a novel multi-agent model-bas…