1 paper
Yuli Chen, Shuhao Zhang, Fanshen Meng +4
Depth pruning improves the deployment efficiency of large language models (LLMs) by identifying and removing redundant layers. A widely accepted standard for this identification pr…