1 paper · 1 filter
Lei Lu, Zhepeng Wang, Runxue Bao +7
Existing pruning techniques for large language models (LLMs) targeting domain-specific applications typically follow a two-stage process: pruning the pretrained general-purpose LLM…