IROS 2018poster14 citations

Policy Shaping with Supervisory Attention Driven Exploration

Taylor Kessler Faulkner, Elaine Schaertl Short, Andrea Lockerd Thomaz

Abstract

Robots deployed for long periods of time need to be able to explore and learn from their environment. One approach to this problem has been reinforcement learning (RL), in which robots receive rewards from the environment that allow them to choose optimal actions. To speed learning when human supervision is available, interactive reinforcement learning solicits feedback from a human teacher. However, this approach typically assumes that learning takes place under continuous supervision, which is unlikely to hold in long-term scenarios. We propose an extension to a method of interactive reinforcement learning, policy shaping, that takes into account human attention. Our approach enables better performance while unattended by favoring information-gathering actions when attended and actions that have received positive feedback when unattended. We test our approach in both simulation and on a robot, finding that our method learns faster than policy shaping and performs more safely than policy shaping while no one is paying attention to the robot.

BibTeX
@inproceedings{iros2018_policyshapingwit,
  title = {Policy Shaping with Supervisory Attention Driven Exploration},
  author = {Taylor Kessler Faulkner and Elaine Schaertl Short and Andrea Lockerd Thomaz},
  booktitle = {IROS 2018},
  year = {2018}
}
Policy Shaping with Supervisory Attention Driven Exploration · IROS 2018