44 citations · 61 across the 28 of their papers we have counts for
1 paper · 2 filters
Oscar Smee, Fred Roosta, Stephen J. Wright
Gradient descent is the primary workhorse for optimizing large-scale problems in machine learning. However, its performance is highly sensitive to the choice of the learning rate.…