1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.AR2022★ 1 cited
S4: a High-sparsity, High-performance AI Accelerator
Ian En-Hsu Yen, Zhibin Xiao, Dongkuan Xu
Exploiting sparsity underlying neural networks has become one of the most potential methodologies to reduce the memory footprint, I/O cost, and computation workloads during inferen…
cs.CL2021
Rethinking Network Pruning -- under the Pre-train and Fine-tune Paradigm
Dongkuan Xu, Ian E. H. Yen, Jinxi Zhao +1
Transformer-based pre-trained language models have significantly improved the performance of various natural language processing (NLP) tasks in the recent years. While effective an…