1 paper · 1 filter
Hang Guo, Yawei Li, Luca Benini
Recent advances in Large Language Model (LLM) compression, such as quantization and pruning, have achieved notable success. However, as these techniques gradually approach their re…