2 papers
cs.DC2019
Analysis of Large-Scale Multi-Tenant GPU Clusters for DNN Training Workloads
Myeongjae Jeon, Shivaram Venkataraman, Amar Phanishayee +3
With widespread advances in machine learning, a number of large enterprises are beginning to incorporate machine learning models across a number of products. These models are typic…
cs.CV2018
Balanced Sparsity for Efficient DNN Inference on GPU
Zhuliang Yao, Shijie Cao, Wencong Xiao +2
In trained deep neural networks, unstructured pruning can reduce redundant weights to lower storage cost. However, it requires the customization of hardwares to speed up practical…