9 citations · 43 across the 22 of their papers we have counts for
Showing cs.DCShow all
2 papers · 1 filter
cs.DC2022★ 17 cited
Efficient Data-Plane Memory Scheduling for In-Network Aggregation
Hao Wang, Yuxuan Qin, ChonLam Lao +3
As the scale of distributed training grows, communication becomes a bottleneck. To accelerate the communication, recent works introduce In-Network Aggregation (INA), which moves th…
cs.DC2021★ 3 cited
Automatic Configuration for Optimal Communication Scheduling in DNN Training
Yiqing Ma, Hao Wang, Yiming Zhang +1
ByteScheduler partitions and rearranges tensor transmissions to improve the communication efficiency of distributed Deep Neural Network (DNN) training. The configuration of hyper-p…