1 paper · 1 filter
Chongyu Qu, Ritchie Zhao, Ye Yu +6
Quantizing deep neural networks ,reducing the precision (bit-width) of their computations, can remarkably decrease memory usage and accelerate processing, making these models more…