2 citations · 4 across the 3 of their papers we have counts for
Showing 2018Show all
3 papers · 1 filter
stat.ML2018
Stagewise Training Accelerates Convergence of Testing Error Over SGD
Zhuoning Yuan, Yan Yan, Rong Jin +1
Stagewise training strategy is widely used for learning neural networks, which runs a stochastic algorithm (e.g., SGD) starting with a relatively large step size (aka learning rate…
cs.LG2018
A Unified Analysis of Stochastic Momentum Methods for Deep Learning
Yan Yan, Tianbao Yang, Zhe Li +2
Stochastic momentum methods have been widely adopted in training deep neural networks. However, their theoretical analysis of convergence of the training objective and the generali…
cs.CV2018
Style Aggregated Network for Facial Landmark Detection
Xuanyi Dong, Yan Yan, Wanli Ouyang +1
Recent advances in facial landmark detection achieve success by learning discriminative features from rich deformation of face shapes and poses. Besides the variance of faces thems…