A Study of Deep Learning Robustness Against Computation Failures
arXiv:1704.05396
Abstract
For many types of integrated circuits, accepting larger failure rates in computations can be used to improve energy efficiency. We study the performance of faulty implementations of certain deep neural networks based on pessimistic and optimistic models of the effect of hardware faults. After identifying the impact of hyperparameters such as the number of layers on robustness, we study the ability of the network to compensate for computational failures through an increase of the network size. We show that some networks can achieve equivalent performance under faulty implementations, and quantify the required increase in computational complexity.
References in corpus (4)
- SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size
- ADADELTA: An Adaptive Learning Rate Method
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin
- Binarized Neural Networks: Training Deep Neural Networks with Weights and Activations Constrained to +1 or -1