2019
Robust exploration in linear quadratic reinforcement learning
NeurIPS 2019spotlight
Learning to make decisions in an uncertain and dynamic environment is a task of fundamental performance in a number of domains. This paper concerns the problem of learning control policies for an unknown linear dynamical system so as to minimize a quadratic cost function. We present a method, based…