Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
Network Pruning Spaces
Xuanyu He, Yu-I Yang, Ran Song +5
Network pruning techniques, including weight pruning and filter pruning, reveal that most state-of-the-art neural networks can be accelerated without a significant performance drop…
cs.CV2020
Accelerating Neural Network Inference by Overflow Aware Quantization
Hongwei Xie, Shuo Zhang, Huanghao Ding +5
The inherent heavy computation of deep neural networks prevents their widespread applications. A widely used method for accelerating model inference is quantization, by replacing t…