211 citations · 475 across the 13 of their papers we have counts for
Showing cs.NEShow all
2 papers · 1 filter
cs.NE2019★ 26 cited
Progressive DNN Compression: A Key to Achieve Ultra-High Weight Pruning and Quantization Rates using ADMM
Shaokai Ye, Xiaoyu Feng, Tianyun Zhang +11
Weight pruning and weight quantization are two important categories of DNN model compression. Prior work on these techniques are mainly based on heuristics. A recent work developed…
cs.NE2018
A Unified Framework of DNN Weight Pruning and Weight Clustering/Quantization Using ADMM
Shaokai Ye, Tianyun Zhang, Kaiqi Zhang +6
Many model compression techniques of Deep Neural Networks (DNNs) have been investigated, including weight pruning, weight clustering and quantization, etc. Weight pruning leverages…