33 citations · 54 across the 18 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
LoaQ: Layer-wise Output Approximation Quantization
Li Lin, Xiaojun Wan
A natural and intuitive idea in model quantization is to approximate each component's quantized output to match its original. Motivated by this idea, most layer-wise post-training…
cs.LG2025
NeUQI: Near-Optimal Uniform Quantization Parameter Initialization for Low-Bit LLMs
Li Lin, Xinyu Hu, Xiaojun Wan
Large language models (LLMs) achieve impressive performance across domains but face significant challenges when deployed on consumer-grade GPUs or personal devices such as laptops,…