Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Sparse Layer Sharpness-Aware Minimization for Efficient Fine-Tuning
Yifei Cheng, Xianglin Yang, Guoxia Wang +5
Sharpness-aware minimization (SAM) seeks the minima with a flat loss landscape to improve the generalization performance in machine learning tasks, including fine-tuning. However,…
cs.LG2024
Neural Surveillance: Live-Update Visualization of Latent Training Dynamics
Xianglin Yang, Jin Song Dong
Monitoring the inner state of deep neural networks is essential for auditing the learning process and enabling timely interventions. While conventional metrics like validation loss…