1 paper · 1 filter
Yuli Chen, Shuhao Zhang, Fanshen Meng +4
Depth pruning improves the deployment efficiency of large language models (LLMs) by identifying and removing redundant layers. A widely accepted standard for this identification pr…