1 paper
Xiao Zhou, Weizhong Zhang, Hang Xu +1
Weight pruning is an effective technique to reduce the model size and inference time for deep neural networks in real-world deployments. However, since magnitudes and relative impo…