1 paper · 1 filter
Damien Teney, Armand Nicolicioiu, Valentin Hartmann +1
Our understanding of the generalization capabilities of neural networks (NNs) is still incomplete. Prevailing explanations are based on implicit biases of gradient descent (GD) but…