1 paper
Jaehyeon Moon, Dohyung Kim, Junyong Cheon +1
Post-training quantization (PTQ) is an efficient model compression technique that quantizes a pretrained full-precision model using only a small calibration set of unlabeled sample…