1 paper
Banseok Lee, Dongkyu Kim, Youngcheon You +1
The deployment of large language models (LLMs) is frequently hindered by prohibitive memory and computational requirements. While quantization mitigates these bottlenecks, maintain…