1 paper
Geon Park, Jaehong Yoon, Haiyang Zhang +3
Neural network quantization aims to transform high-precision weights and activations of a given neural network into low-precision weights/activations for reduced memory usage and c…