2 papers
cs.LG2025
An Effective Gram Matrix Characterizes Generalization in Deep Networks
Rubing Yang, Pratik Chaudhari
We derive a differential equation that governs the evolution of the generalization gap when a deep network is trained by gradient descent. This differential equation is controlled…
stat.ML2025
Prospective Learning: Learning for a Dynamic Future
Ashwin De Silva, Rahul Ramesh, Rubing Yang +3
In real-world applications, the distribution of the data, and our goals, evolve over time. The prevailing theoretical framework for studying machine learning, namely probably appro…