1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Kento Imaizumi, Hideaki Iiduka
The performance of stochastic gradient descent (SGD), which is the simplest first-order optimizer for training deep neural networks, depends on not only the learning rate but also…