67 citations · 67 across the 1 of their papers we have counts for
1 paper
Shaohuai Shi, Xiaowen Chu, Ka Chun Cheung +1
Distributed stochastic gradient descent (SGD) algorithms are widely deployed in training large-scale deep learning models, while the communication overhead among workers becomes th…