85 citations · 155 across the 9 of their papers we have counts for
Showing cs.DCShow all
2 papers · 1 filter
cs.DC2018
Data-parallel distributed training of very large models beyond GPU capacity
Samuel Matzek, Max Grossman, Minsik Cho +3
GPUs have limited memory and it is difficult to train wide and/or deep models that cause the training process to go out of memory. It is shown in this paper how an open source tool…
cs.DC2017★ 4 cited
PowerAI DDL
Minsik Cho, Ulrich Finkler, Sameer Kumar +3
As deep neural networks become more complex and input datasets grow larger, it can take days or even weeks to train a deep neural network to the desired accuracy. Therefore, distri…