1 paper
Yan Luo, Yangcheng Gao, Zhao Zhang +3
Quantization approximates a deep network model with floating-point numbers by the one with low bit width numbers, in order to accelerate inference and reduce computation. Quantizin…