14 citations · 39 across the 9 of their papers we have counts for
1 paper · 1 filter
Juntang Zhuang, Tommy Tang, Yifan Ding +4
Most popular optimizers for deep learning can be broadly categorized as adaptive methods (e.g. Adam) and accelerated schemes (e.g. stochastic gradient descent (SGD) with momentum).…