1 paper
Shipeng Bai, Jun Chen, Xintian Shen +2
Structured pruning and quantization are promising approaches for reducing the inference time and memory footprint of neural networks. However, most existing methods require the ori…