1 citations · 1 across the 1 of their papers we have counts for
1 paper
Peng Wang, Li Shen, Zerui Tao +3
Averaging iterations of Stochastic Gradient Descent (SGD) have achieved empirical success in training deep learning models, such as Stochastic Weight Averaging (SWA), Exponential M…