2 papers
cs.LG2025
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
Xin Yuan, Siqi Li, Jiateng Wei +7
Pruning is an effective method for compressing Large Language Models, but finding an optimal, non-uniform layer-wise sparsity allocation remains a key challenge. While heuristic me…
cs.CV2024
OvSW: Overcoming Silent Weights for Accurate Binary Neural Networks
Jingyang Xiang, Zuohui Chen, Siqi Li +2
Binary Neural Networks~(BNNs) have been proven to be highly effective for deploying deep neural networks on mobile and embedded platforms. Most existing works focus on minimizing q…