61 citations · 107 across the 6 of their papers we have counts for
1 paper · 1 filter
Zheng Wang, Juncheng B Li, Shuhui Qu +2
Quantization is an effective technique to reduce memory footprint, inference latency, and power consumption of deep learning models. However, existing quantization methods suffer f…