1 paper
Jung Hyun Lee, Seungjae Shin, Vinnam Kim +2
As the rapid scaling of large language models (LLMs) poses significant challenges for deployment on resource-constrained devices, there is growing interest in extremely low-bit qua…