1 citations · 2 across the 9 of their papers we have counts for
1 paper · 1 filter
Haihang Wu, Wei Wang, Tamasha Malepathirana +3
Pruning can be an effective method of compressing large pre-trained models for inference speed acceleration. Previous pruning approaches rely on access to the original training dat…