2015
Nonparametric Bayesian reward segmentation for skill discovery using inverse reinforcement learning
IROS 2015poster
We present a method for segmenting a set of unstructured demonstration trajectories to discover reusable skills using inverse reinforcement learning (IRL). Each skill is characterised by a latent reward function which the demonstrator is assumed to be optimizing. The skill boundaries and the number…