41 citations · 214 across the 22 of their papers we have counts for
Showing 2017Show all
2 papers · 1 filter
cs.LG2017★ 9 cited
Interpolated Policy Gradient: Merging On-Policy and Off-Policy Gradient Estimation for Deep Reinforcement Learning
Shixiang Gu, Timothy Lillicrap, Zoubin Ghahramani +3
Off-policy model-free deep reinforcement learning methods using previously collected data can improve sample efficiency over on-policy policy gradient techniques. On the other hand…
stat.ML2017★ 41 cited
Discriminative k-shot learning using probabilistic models
Matthias Bauer, Mateo Rojas-Carulla, Jakub Bartłomiej Świątkowski +2
This paper introduces a probabilistic framework for k-shot image classification. The goal is to generalise from an initial large-scale classification task to a separate task compri…