Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
Xin Yuan, Siqi Li, Jiateng Wei +7
Pruning is an effective method for compressing Large Language Models, but finding an optimal, non-uniform layer-wise sparsity allocation remains a key challenge. While heuristic me…
cs.LG2024
AutoDFP: Automatic Data-Free Pruning via Channel Similarity Reconstruction
Siqi Li, Jun Chen, Jingyang Xiang +2
Structured pruning methods are developed to bridge the gap between the massive scale of neural networks and the limited hardware resources. Most current structured pruning methods…