5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 5 cited
MVQ:Towards Efficient DNN Compression and Acceleration with Masked Vector Quantization
Shuaiting Li, Chengxuan Wang, Juncan Deng +5
Vector quantization(VQ) is a hardware-friendly DNN compression method that can reduce the storage cost and weight-loading datawidth of hardware accelerators. However, conventional…
cs.LG2024
VQ4ALL: Efficient Neural Network Representation via a Universal Codebook
Juncan Deng, Shuaiting Li, Zeyu Wang +3
The rapid growth of the big neural network models puts forward new requirements for lightweight network representation methods. The traditional methods based on model compression h…