109 citations · 184 across the 3 of their papers we have counts for
Showing cs.DCShow all
3 papers · 1 filter
cs.DC2023★ 38 cited
FlexMoE: Scaling Large-scale Sparse Pre-trained Model Training via Dynamic Device Placement
Xiaonan Nie, Xupeng Miao, Zilong Wang +5
With the increasing data volume, there is a trend of using large-scale pre-trained models to store the knowledge into an enormous number of model parameters. The training of these…
cs.DC2018
Towards Efficient Large-Scale Graph Neural Network Computing
Lingxiao Ma, Zhi Yang, Youshan Miao +4
Recent deep learning models have moved beyond low-dimensional regular grids such as image, video, and speech, to high-dimensional graph-structured data, such as social networks, br…
cs.DC2018
RPC Considered Harmful: Fast Distributed Deep Learning on RDMA
Jilong Xue, Youshan Miao, Cheng Chen +3
Deep learning emerges as an important new resource-intensive workload and has been successfully applied in computer vision, speech, natural language processing, and so on. Distribu…