12 citations · 12 across the 3 of their papers we have counts for
1 paper · 1 filter
Lili Su, Pengkun Yang
We consider training over-parameterized two-layer neural networks with Rectified Linear Unit (ReLU) using gradient descent (GD) method. Inspired by a recent line of work, we study…