1 paper · 1 filter
Yingsong Luo, Ling Chen
Large language models (LLMs) excel in various tasks but face deployment challenges due to hardware constraints. We propose density-aware post-training weight-only quantization (DAQ…