1 paper
Zhuowen Liu, Longkun Hao, Shiyu Feng +3
The rapid growth in the parameter scale of large language models (LLMs) has created a strong demand for efficient compression techniques. As a hardware-agnostic and highly compatib…