5 citations · 6 across the 7 of their papers we have counts for
1 paper · 1 filter
Tao Wang, Junsong Wang, Chang Xu +1
Model quantization is a widely used technique to compress and accelerate deep neural network (DNN) inference, especially when deploying to edge or IoT devices with limited computat…