1 paper
Hongwei Xie, Shuo Zhang, Huanghao Ding +5
The inherent heavy computation of deep neural networks prevents their widespread applications. A widely used method for accelerating model inference is quantization, by replacing t…