49 citations · 79 across the 7 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.LG2023
Sharpness-Aware Minimization and the Edge of Stability
Philip M. Long, Peter L. Bartlett
Recent experiments have shown that, often, when training a neural network with gradient descent (GD) with a step size , the operator norm of the Hessian of the loss grows until…
cs.LG2023★ 49 cited
Prediction, Learning, Uniform Convergence, and Scale-sensitive Dimensions
Peter L. Bartlett, Philip M. Long
We present a new general-purpose algorithm for learning classes of -valued functions in a generalization of the prediction model, and prove a general upper bound on the expe…