3 papers
cs.AR2023
IOPS: An Unified SpMM Accelerator Based on Inner-Outer-Hybrid Product
Wenhao Sun, Wendi Sun, Song Chen +1
Sparse matrix multiplication (SpMM) is widely applied to numerous domains, such as graph processing, machine learning, and data analytics. However, inner product based SpMM induces…
cs.DC2023
BandMap: Application Mapping with Bandwidth Allocation forCoarse-Grained Reconfigurable Array
Xiaobing Ni, Jiaheng Ruan, Mengke Ge +3
This paper proposes an application mapping algorithm, BandMap, for coarse-grained reconfigurable array (CGRA), which allocates the bandwidth in PE array according to the transferri…
cs.AR2023
Bit-balance: Model-Hardware Co-design for Accelerating NNs by Exploiting Bit-level Sparsity
Wenhao Sun, Zhiwei Zou, Deng Liu +3
Bit-serial architectures can handle Neural Networks (NNs) with different weight precisions, achieving higher resource efficiency compared with bit-parallel architectures. Besides,…