1 paper
Kaixin Xu, Alina Hui Xiu Lee, Ziyuan Zhao +3
A popular track of network compression approach is Quantization aware Training (QAT), which accelerates the forward pass during the neural network training and inference. However,…