Showing stat.MLShow all
3 papers · 1 filter
stat.ML2026
Hard labels sampled from sparse targets mislead rotation invariant algorithms
Avrajit Ghosh, Bin Yu, Manfred Warmuth +1
One of the most common machine learning setups is logistic regression. In many classification models, including neural networks, the final prediction is obtained by applying a logi…
stat.ML2025
Variational Learning Finds Flatter Solutions at the Edge of Stability
Avrajit Ghosh, Bai Cong, Rio Yokota +5
Variational Learning (VL) has recently gained popularity for training deep neural networks. Part of its empirical success can be explained by theories such as PAC-Bayes bounds, min…
stat.ML2025
Learning Dynamics of Deep Linear Networks Beyond the Edge of Stability
Avrajit Ghosh, Soo Min Kwon, Rongrong Wang +2
Deep neural networks trained using gradient descent with a fixed learning rate often operate in the regime of "edge of stability" (EOS), where the largest eigenvalue of the He…