1 paper · 1 filter
Mingluo Su, Huan Wang
Pruning is widely recognized as an effective method for reducing the parameters of large language models (LLMs), potentially leading to more efficient deployment and inference. One…