Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
Depth Separation with Multilayer Mean-Field Networks
Yunwei Ren, Mo Zhou, Rong Ge
Depth separation -- why a deeper network is more powerful than a shallower one -- has been a major problem in deep learning theory. Previous results often focus on representation p…
cs.LG2023
Implicit Regularization Leads to Benign Overfitting for Sparse Linear Regression
Mo Zhou, Rong Ge
In deep learning, often the training process finds an interpolator (a solution with 0 training loss), but the test loss is still low. This phenomenon, known as benign overfitting,…