Auto Deep Compression by Reinforcement Learning Based Actor-Critic Structure
arXiv:1807.02886
Abstract
Model-based compression is an effective, facilitating, and expanded model of neural network models with limited computing and low power. However, conventional models of compression techniques utilize crafted features [2,3,12] and explore specialized areas for exploration and design of large spaces in terms of size, speed, and accuracy, which usually have returns Less and time is up. This paper will effectively analyze deep auto compression (ADC) and reinforcement learning strength in an effective sample and space design, and improve the compression quality of the model. The results of compression of the advanced model are obtained without any human effort and in a completely automated way. With a 4- fold reduction in FLOP, the accuracy of 2.8% is higher than the manual compression model for VGG-16 in ImageNet.
References in corpus (17)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
- Neural Architecture Search with Reinforcement Learning
- Compressing Deep Convolutional Networks using Vector Quantization
- ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices
- Trained Ternary Quantization
- Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures
- Pruning Filters for Efficient ConvNets
- Speeding up Convolutional Neural Networks with Low Rank Expansions
- SMASH: One-Shot Model Architecture Search through HyperNetworks
- Channel Pruning for Accelerating Very Deep Neural Networks
- Speeding-up Convolutional Neural Networks Using Fine-tuned CP-Decomposition
- Fast Convolutional Nets With fbfft: A GPU Performance Evaluation
- N2N Learning: Network to Network Compression via Policy Gradient Reinforcement Learning
- More is Less: A More Complicated Network with Less Inference Complexity
- Compact Deep Convolutional Neural Networks With Coarse Pruning
- Practical Block-wise Neural Network Architecture Generation