← Search

Alfredo Canziani

2 accepted papers

2019

Model-Predictive Policy Learning with Uncertainty Regularization for Driving in Dense Traffic

ICLR 2019poster

Learning a policy using only observational data is challenging because the distribution of states it induces at execution time may differ from the distribution observed during training. In this work, we propose to train a policy while explicitly penalizing the mismatch between these two distribution…