116 citations · 122 across the 2 of their papers we have counts for
1 paper · 1 filter
Chen Xing, Devansh Arpit, Christos Tsirigotis +1
We present novel empirical observations regarding how stochastic gradient descent (SGD) navigates the loss landscape of over-parametrized deep neural networks (DNNs). These observa…