← Search

Giuseppe Canonaco

2 accepted papers

2021

Time-variant variational transfer for value functions

UAI 2021poster

In most of the transfer learning approaches to reinforcement learning (RL) the distribution over the tasks is assumed to be stationary. Therefore, the target and source tasks are i.i.d. samples of the same distribution. Unfortunately, this assumption rarely holds in real-world conditions, e.g., due…

2018

Stochastic Variance-Reduced Policy Gradient

ICML 2018oral

In this paper, we propose a novel reinforcement-learning algorithm consisting in a stochastic variance-reduced version of policy gradient for solving Markov Decision Processes (MDPs). Stochastic variance-reduced gradient (SVRG) methods have proven to be very successful in supervised learning. Howeve…