8 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.LG2023
AutoQNN: An End-to-End Framework for Automatically Quantizing Neural Networks
Cheng Gong, Ye Lu, Surong Dai +3
Exploring the expected quantizing scheme with suitable mixed-precision policy is the key point to compress deep neural networks (DNNs) in high efficiency and accuracy. This explora…
cs.CV2021★ 8 cited
Elastic Significant Bit Quantization and Acceleration for Deep Neural Networks
Cheng Gong, Ye Lu, Kunpeng Xie +3
Quantization has been proven to be a vital method for improving the inference efficiency of deep neural networks (DNNs). However, it is still challenging to strike a good balance b…