2 papers
cs.LG2020
Time-Variant Variational Transfer for Value Functions
Giuseppe Canonaco, Andrea Soprani, Manuel Roveri +1
In most of the transfer learning approaches to reinforcement learning (RL) the distribution over the tasks is assumed to be stationary. Therefore, the target and source tasks are i…
cs.LG2018
Stochastic Variance-Reduced Policy Gradient
Matteo Papini, Damiano Binaghi, Giuseppe Canonaco +2
In this paper, we propose a novel reinforcement- learning algorithm consisting in a stochastic variance-reduced version of policy gradient for solving Markov Decision Processes (MD…