← Search

Tom Van de Wiele

2 accepted papers

2020

Fast Task Inference with Variational Intrinsic Successor Features

ICLR 2020talk

It has been established that diverse behaviors spanning the controllable subspace of a Markov decision process can be trained by rewarding a policy for being distinguishable from other policies. However, one limitation of this formulation is the difficulty to generalize beyond the finite set of beha…

Cited by 203SourceScholar
2019

Unsupervised Control Through Non-Parametric Discriminative Rewards

ICLR 2019poster

Learning to control an environment without hand-crafted rewards or expert data remains challenging and is at the frontier of reinforcement learning research. We present an unsupervised learning algorithm to train agents to achieve perceptually-specified goals using only a stream of observations and…

Cited by 201SourcePDFScholar