1 paper · 1 filter
Romain Petit, Clarice Poon, Gabriel Peyré +1
A surprising phenomenon in the training of neural networks is the ability of gradient descent to find global minimizers of the training loss despite its non-convexity. Following ea…