6 citations · 6 across the 1 of their papers we have counts for
1 paper
Mikhail Rudakov, Aleksandr Beznosikov, Yaroslav Kholodov +1
Large neural networks require enormous computational clusters of machines. Model-parallel training, when the model architecture is partitioned sequentially between workers, is a po…