1 citations · 1 across the 1 of their papers we have counts for
1 paper
Zijing Gu
We implemented and optimized matrix multiplications between dense and block-sparse matrices on CUDA. We leveraged TVM, a deep learning compiler, to explore the schedule space of th…