439 citations · 582 across the 2 of their papers we have counts for
2 papers
cs.LG2019★ 439 cited
The State of Sparsity in Deep Neural Networks
Trevor Gale, Erich Elsen, Sara Hooker
We rigorously evaluate three state-of-the-art techniques for inducing sparsity in deep neural networks on two large-scale learning tasks: Transformer trained on WMT 2014 English-to…
cs.CV2016★ 143 cited
DSD: Dense-Sparse-Dense Training for Deep Neural Networks
Song Han, Jeff Pool, Sharan Narang +9
Modern deep neural networks have a large number of parameters, making them very hard to train. We propose DSD, a dense-sparse-dense training flow, for regularizing deep neural netw…