3 papers
cs.LG2026
Sparse Layer Sharpness-Aware Minimization for Efficient Fine-Tuning
Yifei Cheng, Xianglin Yang, Guoxia Wang +5
Sharpness-aware minimization (SAM) seeks the minima with a flat loss landscape to improve the generalization performance in machine learning tasks, including fine-tuning. However,…
cs.LG2025
LightSAM: Parameter-Agnostic Sharpness-Aware Minimization
Yifei Cheng, Li Shen, Hao Sun +3
Sharpness-Aware Minimization (SAM) optimizer enhances the generalization ability of the machine learning model by exploring the flat minima landscape through weight perturbations.…
cs.DC2024
Adaptive Weighting Push-SUM for Decentralized Optimization with Statistical Diversity
Yiming Zhou, Yifei Cheng, Linli Xu +1
Statistical diversity is a property of data distribution and can hinder the optimization of a decentralized network. However, the theoretical limitations of the Push-SUM protocol r…