Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Compensate Quantization Errors+: Quantized Models Are Inquisitive Learners
Yifei Gao, Jie Ou, Lei Wang +2
The quantization of large language models (LLMs) has been a prominent research area aimed at enabling their lightweight deployment in practice. Existing research about LLM's quanti…
cs.CL2024
Compensate Quantization Errors: Make Weights Hierarchical to Compensate Each Other
Yifei Gao, Jie Ou, Lei Wang +4
Emergent Large Language Models (LLMs) use their extraordinary performance and powerful deduction capacity to discern from traditional language models. However, the expenses of comp…