1 paper · 1 filter
Hongxi Li, Chunlin Huang
We present a theory of feature learning in wide L2-regularized networks showing that supervised learning is inherently compressive. We derive a kernel ODE that predicts a "water-fi…