1 paper
Dongyun Kam, Myeongji Yun, Sunwoo Yoo +3
Low bit-precisions and their bit-slice sparsity have recently been studied to accelerate general matrix-multiplications (GEMM) during large-scale deep neural network (DNN) inferenc…