2017
Dynamic Safe Interruptibility for Decentralized Multi-Agent Reinforcement Learning
NeurIPS 2017spotlight
In reinforcement learning, agents learn by performing actions and observing their outcomes. Sometimes, it is desirable for a human operator to interrupt an agent in order to prevent dangerous situations from happening. Yet, as part of their learning process, agents may link these interruptions, that…