7 citations · 7 across the 1 of their papers we have counts for
Showing cs.DCShow all
2 papers · 1 filter
cs.DC2019
Achieving Super-Linear Speedup across Multi-FPGA for Real-Time DNN Inference
Weiwen Jiang, Edwin H. -M. Sha, Xinyi Zhang +4
Real-time Deep Neural Network (DNN) inference with low-latency requirement has become increasingly important for numerous applications in both cloud computing (e.g., Apple's Siri)…
cs.DC2019★ 7 cited
Accuracy vs. Efficiency: Achieving Both through FPGA-Implementation Aware Neural Architecture Search
Weiwen Jiang, Xinyi Zhang, Edwin H. -M. Sha +4
A fundamental question lies in almost every application of deep neural networks: what is the optimal neural architecture given a specific dataset? Recently, several Neural Architec…