393 citations · 468 across the 2 of their papers we have counts for
1 paper · 1 filter
Andrew Brock, Theodore Lim, J. M. Ritchie +1
The early layers of a deep neural net have the fewest parameters, but take up the most computation. In this extended abstract, we propose to only train the hidden layers for a set…