1 paper
Zhirui Hu, Peiyan Dong, Zhepeng Wang +3
Model compression, such as pruning and quantization, has been widely applied to optimize neural networks on resource-limited classical devices. Recently, there are growing interest…