6 citations · 10 across the 2 of their papers we have counts for
3 papers
cs.PF2020★ 6 cited
Optimizing Memory-Access Patterns for Deep Learning Accelerators
Hongbin Zheng, Sejong Oh, Huiqing Wang +7
Deep learning (DL) workloads are moving towards accelerators for faster processing and lower cost. Modern DL accelerators are good at handling the large-scale multiply-accumulate o…
cs.DC2019★ 4 cited
A Unified Optimization Approach for CNN Model Inference on Integrated GPUs
Leyuan Wang, Zhi Chen, Yizhi Liu +4
Modern deep learning applications urge to push the model inference taking place at the edge devices for multiple reasons such as achieving shorter latency, relieving the burden of…
cs.DC2018
Optimizing CNN Model Inference on CPUs
Yizhi Liu, Yao Wang, Ruofei Yu +3
The popularity of Convolutional Neural Network (CNN) models and the ubiquity of CPUs imply that better performance of CNN model inference on CPUs can deliver significant gain to a…