15 citations · 15 across the 1 of their papers we have counts for
1 paper
Negar Foroutan Eghlidi, Martin Jaggi
Synchronous stochastic gradient descent (SGD) is the most common method used for distributed training of deep learning models. In this algorithm, each worker shares its local gradi…