117 citations · 240 across the 13 of their papers we have counts for
Showing 2020 · cs.LGShow all
2 papers · 2 filters
cs.LG2020★ 10 cited
Wide flat minima and optimal generalization in classifying high-dimensional Gaussian mixtures
Carlo Baldassi, Enrico M. Malatesta, Matteo Negri +1
We analyze the connection between minimizers with good generalizing properties and high local entropy regions of a threshold-linear classifier in Gaussian mixtures with the mean sq…
cs.LG2020
Entropic gradient descent algorithms and wide flat minima
Fabrizio Pittorino, Carlo Lucibello, Christoph Feinauer +4
The properties of flat minima in the empirical risk landscape of neural networks have been debated for some time. Increasing evidence suggests they possess better generalization ca…