171 citations · 289 across the 15 of their papers we have counts for
1 paper · 1 filter
Yuji Chai, John Gkountouras, Glenn G. Ko +2
We introduce a method that dramatically reduces fine-tuning VRAM requirements and rectifies quantization errors in quantized Large Language Models. First, we develop an extremely m…