2k citations · 2k across the 1 of their papers we have counts for
1 paper
Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou +3
While depth tends to improve network performances, it also makes gradient-based training more difficult since deeper networks tend to be more non-linear. The recently proposed know…