22 citations · 96 across the 10 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2021★ 4 cited
How to Design Sample and Computationally Efficient VQA Models
Karan Samel, Zelin Zhao, Binghong Chen +3
In multi-modal reasoning tasks, such as visual question answering (VQA), there have been many modeling and training paradigms tested. Previous models propose different methods for…
cs.LG2020★ 22 cited
APQ: Joint Search for Network Architecture, Pruning and Quantization Policy
Tianzhe Wang, Kuan Wang, Han Cai +3
We present APQ for efficient deep learning inference on resource-constrained hardware. Unlike previous methods that separately search the neural architecture, pruning policy, and q…
cs.LG2019★ 17 cited
Design Automation for Efficient Deep Learning Computing
Song Han, Han Cai, Ligeng Zhu +4
Efficient deep learning computing requires algorithm and hardware co-design to enable specialization: we usually need to change the algorithm to reduce memory footprint and improve…