3 papers
cs.AR2021
Virtual-Link: A Scalable Multi-Producer, Multi-Consumer Message Queue Architecture for Cross-Core Communication
Qinzhe Wu, Jonathan Beard, Ashen Ekanayake +2
Cross-core communication is increasingly a bottleneck as the number of processing elements increase per system-on-chip. Typical hardware solutions to cross-core communication are o…
cs.LG2019
Demystifying the MLPerf Benchmark Suite
Snehil Verma, Qinzhe Wu, Bagus Hanindhito +4
MLPerf, an emerging machine learning benchmark suite strives to cover a broad range of applications of machine learning. We present a study on its characteristics and how the MLPer…
cs.DC2018
Start Late or Finish Early: A Distributed Graph Processing System with Redundancy Reduction
Shuang Song, Xu Liu, Qinzhe Wu +3
Graph processing systems are important in the big data domain. However, processing graphs in parallel often introduces redundant computations in existing algorithms and models. Pri…