2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 2 cited
Hindsight Value Function for Variance Reduction in Stochastic Dynamic Environment
Jiaming Guo, Rui Zhang, Xishan Zhang +6
Policy gradient methods are appealing in deep reinforcement learning but suffer from high variance of gradient estimate. To reduce the variance, the state value function is applied…
cs.IR2021
Dual Adversarial Variational Embedding for Robust Recommendation
Qiaomin Yi, Ning Yang, Philip S. Yu
Robust recommendation aims at capturing true preference of users from noisy data, for which there are two lines of methods have been proposed. One is based on noise injection, and…