2020
SVRG for Policy Evaluation with Fewer Gradient Evaluations
IJCAI 2020poster
Stochastic variance-reduced gradient (SVRG) is an optimization method originally designed for tackling machine learning problems with a finite sum structure. SVRG was later shown to work for policy evaluation, a problem in reinforcement learning in which one aims to estimate the value function of a…