8 citations · 8 across the 2 of their papers we have counts for
1 paper · 1 filter
Kwangmin Yu, Thomas Flynn, Shinjae Yoo +1
Stochastic Gradient Descent (SGD) is the most popular algorithm for training deep neural networks (DNNs). As larger networks and datasets cause longer training times, training on d…