148 citations · 174 across the 8 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
AdaSAM: Boosting Sharpness-Aware Minimization with Adaptive Learning Rate and Momentum for Training Deep Neural Networks
Hao Sun, Li Shen, Qihuang Zhong +6
Sharpness aware minimization (SAM) optimizer has been extensively explored as it can generalize better for training deep neural networks via introducing extra perturbation steps to…
cs.LG2022★ 3 cited
Comprehensive Graph Gradual Pruning for Sparse Training in Graph Neural Networks
Chuang Liu, Xueqi Ma, Yibing Zhan +5
Graph Neural Networks (GNNs) tend to suffer from high computation costs due to the exponentially increasing scale of graph data and the number of model parameters, which restricts…