1 paper
Pietro Tropeano, Maria Maistro, Tuukka Ruotsalo +1
Language Model (LM) pruning compresses the model by removing weights, nodes, or other parts of its architecture. Typically, pruning focuses on the resulting efficiency gains at the…