1 paper
Ryan Lucas, Mehdi Makni, Xiang Meng +2
Quantization is an effective strategy to reduce the storage and computation footprint of large language models (LLMs). Post-training quantization (PTQ) is a leading approach for co…