1 paper
Alex Buna, Shirley Xiaoqi Liu, Patrick Rebeschini
In overparameterised classification, training data can be linearly separable even when the underlying distribution is not. In this setting, gradient descent (GD) on the logistic lo…