← Search

Jonathan W Lavington

1 accepted papers

2021

Robust Asymmetric Learning in POMDPs

ICML 2021oral

Policies for partially observed Markov decision processes can be efficiently learned by imitating expert policies generated using asymmetric information. Unfortunately, existing approaches for this kind of imitation learning have a serious flaw: the expert does not know what the trainee cannot see,…