12 citations · 16 across the 3 of their papers we have counts for
4 papers
Towards Latency-aware DNN Optimization with GPU Runtime Analysis and Tail Effect Elimination
Fuxun Yu, Zirui Xu, Tong Shen +12
Despite the superb performance of State-Of-The-Art (SOTA) DNNs, the increasing computational cost makes them very challenging to meet real-time latency and accuracy requirements. A…
Third ArchEdge Workshop: Exploring the Design Space of Efficient Deep Neural Networks
Fuxun Yu, Dimitrios Stamoulis, Di Wang +2
This paper gives an overview of our ongoing work on the design space exploration of efficient deep neural networks (DNNs). Specifically, we cover two aspects: (1) static architectu…
Single-Path NAS: Device-Aware Efficient ConvNet Design
Dimitrios Stamoulis, Ruizhou Ding, Di Wang +4
Can we automatically design a Convolutional Network (ConvNet) with the highest image classification accuracy under the latency constraint of a mobile device? Neural Architecture Se…
Single-Path NAS: Designing Hardware-Efficient ConvNets in less than 4 Hours
Dimitrios Stamoulis, Ruizhou Ding, Di Wang +4
Can we automatically design a Convolutional Network (ConvNet) with the highest image classification accuracy under the runtime constraint of a mobile device? Neural architecture se…