1 paper
Seungmin Yu, Xiaodie Yi, Hayun Lee +1
N:M sparsity pruning is a powerful technique for compressing deep neural networks, utilizing NVIDIA's Sparse Tensor Core technology. This method benefits from hardware support for…