← Search

Zilun Peng

1 accepted papers

2020

SVRG for Policy Evaluation with Fewer Gradient Evaluations

IJCAI 2020poster

Stochastic variance-reduced gradient (SVRG) is an optimization method originally designed for tackling machine learning problems with a finite sum structure. SVRG was later shown to work for policy evaluation, a problem in reinforcement learning in which one aims to estimate the value function of a…

Cited by 0SourcePDFScholar