1 paper · 1 filter
Donghyun Lee, Yuhang Li, Ruokai Yin +1
Post-training quantization (PTQ) is a widely adopted technique for compressing large language models (LLMs) without retraining. Most existing second-order PTQ methods, including GP…