1 paper
Zhihang Yuan, Yuzhang Shang, Yue Song +4
In this paper, we introduce a new post-training compression paradigm for Large Language Models (LLMs) to facilitate their wider adoption. We delve into LLM weight low-rank decompos…