42 citations · 63 across the 3 of their papers we have counts for
3 papers
cs.LG2019★ 4 cited
Global Convergence of Gradient Descent for Deep Linear Residual Networks
Lei Wu, Qingcan Wang, Chao Ma
We analyze the global convergence of gradient descent for deep linear residual networks by proposing a new initialization: zero-asymmetric (ZAS) initialization. It is motivated by…
cs.LG2019★ 17 cited
Analysis of the Gradient Descent Algorithm for a Deep Neural Network Model with Skip-connections
Weinan E, Chao Ma, Qingcan Wang +1
The behavior of the gradient descent (GD) algorithm is analyzed for a deep neural network model with skip-connections. It is proved that in the over-parametrized regime, for a suit…
cs.LG2019★ 42 cited
A Priori Estimates of the Population Risk for Residual Networks
Weinan E, Chao Ma, Qingcan Wang
Optimal a priori estimates are derived for the population risk, also known as the generalization error, of a regularized residual network model. An important part of the regularize…