1 paper · 1 filter
Arthur Negrão, Pedro Silva, Vander L. S. Freitas +2
The rapid growth of Large Language Models (LLMs) intensifies the need for effective compression, with weight quantization being the most widely adopted technique. Standard uniform…