1 paper
Dohyung Kim, Junghyup Lee, Jeimin Jeon +2
Network quantization generally converts full-precision weights and/or activations into low-bit fixed-point values in order to accelerate an inference process. Recent approaches to…