12 citations · 12 across the 3 of their papers we have counts for
1 paper · 1 filter
Shaohuai Shi, Lin Zhang, Bo Li
Distributed training with synchronous stochastic gradient descent (SGD) on GPU clusters has been widely used to accelerate the training process of deep models. However, SGD only ut…