1 paper
Yuhan Kang, Yang Shi, Mei We +5
Post-training pruning, as one of the key techniques for compressing large language models, plays a vital role in lightweight model deployment and model sparsity. However, current m…