2 papers
cs.LG2024
Forget Sharpness: Perturbed Forgetting of Model Biases Within SAM Dynamics
Ankit Vani, Frederick Tung, Gabriel L. Oliveira +1
Despite attaining high empirical generalization, the sharpness of models trained with sharpness-aware minimization (SAM) do not always correlate with generalization error. Instead…
cs.LG2023
AdaFlood: Adaptive Flood Regularization
Wonho Bae, Yi Ren, Mohamad Osama Ahmed +3
Although neural networks are conventionally optimized towards zero training loss, it has been recently learned that targeting a non-zero training loss threshold, referred to as a f…