ESPN: Extremely Sparse Pruned Networks
arXiv:2006.15741
Abstract
Deep neural networks are often highly overparameterized, prohibiting their use in compute-limited systems. However, a line of recent works has shown that the size of deep networks can be considerably reduced by identifying a subset of neuron indicators (or mask) that correspond to significant weights prior to training. We demonstrate that an simple iterative mask discovery method can achieve state-of-the-art compression of very deep networks. Our algorithm represents a hybrid approach between single shot network pruning methods (such as SNIP) with Lottery-Ticket type approaches. We validate our approach on several datasets and outperform several existing pruning approaches in both test accuracy and compression ratio.
References in corpus (7)
- Quantized Neural Networks: Training Neural Networks with Low Precision Weights and Activations
- Compressing Neural Networks with the Hashing Trick
- Comparing Rewinding and Fine-tuning in Neural Network Pruning
- Memory Bounded Deep Convolutional Networks
- Parameter Efficient Training of Deep Convolutional Neural Networks by Dynamic Sparse Reparameterization
- Deep Model Compression: Distilling Knowledge from Noisy Teachers
- Pruning via Iterative Ranking of Sensitivity Statistics