1 paper
Yeonsik Park, Hyeonseong Kim, Seungkyu Choi
Post-training quantization (PTQ) has emerged as a prevailing technique for deploying large language models (LLMs) efficiently in terms of both memory and computation, across edge d…