1 paper · 1 filter
Dongmin Lee, William Lu, Anuran Makur
Recent work on first-order optimizers for empirical risk minimization (ERM) has suggested that smoothness of ERM loss functions in the training data, rather than in the optimizatio…