1 paper
Zechen Ma, Zixi Qu, Jinyan Yi +2
Distributed machine learning (ML) training has become a necessity with the prevalence of billion to trillion-parameter-scale models. While prior work has improved training efficien…