127 citations · 278 across the 14 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2020★ 18 cited
Experience Replay with Likelihood-free Importance Weights
Samarth Sinha, Jiaming Song, Animesh Garg +1
The use of past experiences to accelerate temporal difference (TD) learning of value functions, or experience replay, is a key component in deep reinforcement learning. Prioritizat…
cs.AI2018★ 2 cited
An Empirical Analysis of Proximal Policy Optimization with Kronecker-factored Natural Gradients
Jiaming Song, Yuhuai Wu
In this technical report, we consider an approach that combines the PPO objective and K-FAC natural gradient optimization, for which we call PPOKFAC. We perform a range of empirica…