1 paper · 1 filter
Junhao Xia, Ming Zhao, Limin Xiao +1
Large language models (LLMs) face significant computational and memory challenges, making extremely low-bit quantization crucial for their efficient deployment. In this work, we in…