1 paper
Fu-Ming Guo, Austin Huang
Reducing computation cost, inference latency, and memory footprint of neural networks are frequently cited as research motivations for pruning and sparsity. However, operationalizi…