5 citations · 5 across the 1 of their papers we have counts for
2 papers
cs.AI2021★ 5 cited
Characterizing the Gap Between Actor-Critic and Policy Gradient
Junfeng Wen, Saurabh Kumar, Ramki Gummadi +1
Actor-critic (AC) methods are ubiquitous in reinforcement learning. Although it is understood that AC methods are closely related to policy gradient (PG), their precise connection…
stat.ML2018
Variational Rejection Sampling
Aditya Grover, Ramki Gummadi, Miguel Lazaro-Gredilla +2
Learning latent variable models with stochastic variational inference is challenging when the approximate posterior is far from the true posterior, due to high variance in the grad…