1 paper
Lianbo Ma, Yuee Zhou, Jianlun Ma +2
Weight quantization is an effective technique to compress deep neural networks for their deployment on edge devices with limited resources. Traditional loss-aware quantization meth…