4 citations · 9 across the 4 of their papers we have counts for
3 papers · 1 filter
Efficient Training of Convolutional Neural Nets on Large Distributed Systems
Sameer Kumar, Dheeraj Sreedhar, Vaibhav Saxena +2
Deep Neural Networks (DNNs) have achieved im- pressive accuracy in many application domains including im- age classification. Training of DNNs is an extremely compute- intensive pr…
PowerAI DDL
Minsik Cho, Ulrich Finkler, Sameer Kumar +3
As deep neural networks become more complex and input datasets grow larger, it can take days or even weeks to train a deep neural network to the desired accuracy. Therefore, distri…
On Optimizing Distributed Tucker Decomposition for Dense Tensors
Venkatesan T Chakaravarthy, Jee W Choi, Douglas J Joseph +4
The Tucker decomposition expresses a given tensor as the product of a small core tensor and a set of factor matrices. Apart from providing data compression, the construction is use…