2 papers
cs.DC2021
End-to-end Adaptive Distributed Training on PaddlePaddle
Yulong Ao, Zhihua Wu, Dianhai Yu +10
Distributed training has become a pervasive and effective approach for training a large neural network (NN) model with processing massive data. However, it is very challenging to s…
cs.DC2020
Adaptive SpMV/SpMSpV on GPUs for Input Vectors of Varied Sparsity
Min Li, Yulong Ao, Chao Yang
Despite numerous efforts for optimizing the performance of Sparse Matrix and Vector Multiplication (SpMV) on modern hardware architectures, few works are done to its sparse counter…