1 paper
Lei Lu, Zhepeng Wang, Runxue Bao +7
Existing pruning techniques for large language models (LLMs) targeting domain-specific applications typically follow a two-stage process: pruning the pretrained general-purpose LLM…