2 papers
cs.LG2024
Sharpness-Aware Minimization Revisited: Weighted Sharpness as a Regularization Term
Yun Yue, Jiadi Jiang, Zhiling Ye +3
Deep Neural Networks (DNNs) generalization is known to be closely related to the flatness of minima, leading to the development of Sharpness-Aware Minimization (SAM) for seeking fl…
cs.LG2024
Adaptive Optimizers with Sparse Group Lasso for Neural Networks in CTR Prediction
Yun Yue, Yongchao Liu, Suo Tong +7
We develop a novel framework that adds the regularizers of the sparse group lasso to a family of adaptive optimizers in deep learning, such as Momentum, Adagrad, Adam, AMSGrad, Ada…