1 paper · 1 filter
Yao Fu, Runchao Li, Xianxuan Long +4
Neural network pruning has emerged as a promising approach for deploying LLMs in low-resource scenarios while preserving downstream task performance. However, for the first time, w…