11 citations · 15 across the 3 of their papers we have counts for
Showing 2020Show all
3 papers · 1 filter
cs.AR2020★ 4 cited
Towards Latency-aware DNN Optimization with GPU Runtime Analysis and Tail Effect Elimination
Fuxun Yu, Zirui Xu, Tong Shen +12
Despite the superb performance of State-Of-The-Art (SOTA) DNNs, the increasing computational cost makes them very challenging to meet real-time latency and accuracy requirements. A…
cs.CV2020
AntiDote: Attention-based Dynamic Optimization for Neural Network Runtime Efficiency
Fuxun Yu, Chenchen Liu, Di Wang +2
Convolutional Neural Networks (CNNs) achieved great cognitive performance at the expense of considerable computation load. To relieve the computation load, many optimization works…
cs.CV2020
An Image Enhancing Pattern-based Sparsity for Real-time Inference on Mobile Devices
Xiaolong Ma, Wei Niu, Tianyun Zhang +8
Weight pruning has been widely acknowledged as a straightforward and effective method to eliminate redundancy in Deep Neural Networks (DNN), thereby achieving acceleration on vario…