4 citations · 4 across the 1 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020★ 4 cited
Variance Reduction for Deep Q-Learning using Stochastic Recursive Gradient
Haonan Jia, Xiao Zhang, Jun Xu +4
Deep Q-learning algorithms often suffer from poor gradient estimations with an excessive variance, resulting in unstable training and poor sampling efficiency. Stochastic variance-…
cs.LG2018
MQGrad: Reinforcement Learning of Gradient Quantization in Parameter Server
Guoxin Cui, Jun Xu, Wei Zeng +3
One of the most significant bottleneck in training large scale machine learning models on parameter server (PS) is the communication overhead, because it needs to frequently exchan…