1 paper
Ian Colbert, Alessandro Pappalardo, Jakoba Petri-Koenig +1
Quantization techniques commonly reduce the inference costs of neural networks by restricting the precision of weights and activations. Recent studies show that also reducing the p…