1 citations · 1 across the 1 of their papers we have counts for
1 paper
Xing Hu, Yuan Cheng, Dawei Yang +4
Post-training quantization (PTQ) serves as a potent technique to accelerate the inference of large language models (LLMs). Nonetheless, existing works still necessitate a considera…