2 citations · 2 across the 1 of their papers we have counts for
1 paper
Sanjith Athlur, Nitika Saran, Muthian Sivathanu +2
Systems for training massive deep learning models (billions of parameters) today assume and require specialized "hyper-clusters": hundreds or thousands of GPUs wired with specializ…