2020
GenDICE: Generalized Offline Estimation of Stationary Values
ICLR 2020talk
An important problem that arises in reinforcement learning and Monte Carlo methods is estimating quantities defined by the stationary distribution of a Markov chain. In many real-world applications, access to the underlying transition operator is limited to a fixed set of data that has already been…