← Search

Damiano Binaghi

1 accepted papers

2018

Stochastic Variance-Reduced Policy Gradient

ICML 2018oral

In this paper, we propose a novel reinforcement-learning algorithm consisting in a stochastic variance-reduced version of policy gradient for solving Markov Decision Processes (MDPs). Stochastic variance-reduced gradient (SVRG) methods have proven to be very successful in supervised learning. Howeve…