7 citations · 11 across the 2 of their papers we have counts for
1 paper · 1 filter
Xinyu Zhang, Ian Colbert, Ken Kreutz-Delgado +1
Quantization and pruning are core techniques used to reduce the inference costs of deep neural networks. State-of-the-art quantization techniques are currently applied to both the…