609 citations · 1k across the 4 of their papers we have counts for
1 paper · 1 filter
Xinghao Pan, Jianmin Chen, Rajat Monga +2
Distributed training of deep learning models on large-scale training data is typically conducted with asynchronous stochastic optimization to maximize the rate of updates, at the c…