← Search

Remi Tachet Combes

1 accepted papers

2020

A Reduction from Reinforcement Learning to No-Regret Online Learning

AISTATS 2020poster

We present a reduction from reinforcement learning (RL) to no-regret online learning based on the saddle-point formulation of RL, by which "any" online algorithm with sublinear regret can generate policies with provable performance guarantees. This new perspective decouples the RL problem into two p…

Cited by 18SourcePDFScholar