6 citations · 12 across the 5 of their papers we have counts for
1 paper · 1 filter
Charles G. Frye, James Simon, Neha S. Wadia +3
Despite the fact that the loss functions of deep neural networks are highly non-convex, gradient-based optimization algorithms converge to approximately the same performance from m…