1 citations · 1 across the 1 of their papers we have counts for
1 paper
Tianyi Chen, Ziye Guo, Yuejiao Sun +1
Stochastic gradient descent (SGD) has taken the stage as the primary workhorse for large-scale machine learning. It is often used with its adaptive variants such as AdaGrad, Adam,…