341 citations · 371 across the 27 of their papers we have counts for
1 paper · 1 filter
Eldar Kurtic, Dan Alistarh
We revisit the performance of the classic gradual magnitude pruning (GMP) baseline for large language models, focusing on the classic BERT benchmark on various popular tasks. Despi…