2 papers
cs.LG2025
When Bias Meets Trainability: Connecting Theories of Initialization
Alberto Bassi, Marco Baity-Jesi, Aurelien Lucchi +2
The statistical properties of deep neural networks (DNNs) at initialization play an important role to comprehend their trainability and the intrinsic architectural biases they poss…
cs.LG2025
Where You Place the Norm Matters: From Prejudiced to Neutral Initializations
Emanuele Francazi, Francesco Pinto, Aurelien Lucchi +1
Normalization layers were introduced to stabilize and accelerate training, yet their influence is critical already at initialization, where they shape signal propagation and output…