3 papers
cs.LG2025
A Theory of Initialisation's Impact on Specialisation
Devon Jarvis, Sebastian Lee, Clémentine Carla Juliette Dominé +2
Prior work has demonstrated a consistent tendency in neural networks engaged in continual learning tasks, wherein intermediate task similarity results in the highest levels of cata…
stat.ML2025
Position: Solve Layerwise Linear Models First to Understand Neural Dynamical Phenomena (Neural Collapse, Emergence, Lazy/Rich Regime, and Grokking)
Yoonsoo Nam, Seok Hyeong Lee, Clementine C J Domine +5
In physics, complex systems are often simplified into minimal, solvable models that retain only the core principles. In machine learning, layerwise linear models (e.g., linear neur…
cs.LG2024
From Lazy to Rich: Exact Learning Dynamics in Deep Linear Networks
Clémentine C. J. Dominé, Nicolas Anguita, Alexandra M. Proca +4
Biological and artificial neural networks develop internal representations that enable them to perform complex tasks. In artificial networks, the effectiveness of these models reli…