5 citations · 7 across the 2 of their papers we have counts for
1 paper · 1 filter
Zihao Deng, Sayeh Sharify, Xin Wang +1
Quantization is a widely used technique to compress neural networks. Assigning uniform bit-widths across all layers can result in significant accuracy degradation at low precision…