13 citations · 24 across the 6 of their papers we have counts for
1 paper · 1 filter
Dmitrii Marin, Meng Tang, Ismail Ben Ayed +1
The simplicity of gradient descent (GD) made it the default method for training ever-deeper and complex neural networks. Both loss functions and architectures are often explicitly…