18 citations · 150 across the 33 of their papers we have counts for
3 papers · 2 filters
Benchmarking network fabrics for data distributed training of deep neural networks
Siddharth Samsi, Andrew Prout, Michael Jones +16
Artificial Intelligence/Machine Learning applications require the training of complex models on large amounts of labelled data. The large computational requirements for training de…
Best of Both Worlds: High Performance Interactive and Batch Launching
Chansup Byun, Jeremy Kepner, William Arcand +16
Rapid launch of thousands of jobs is essential for effective interactive supercomputing, big data analysis, and AI algorithm development. Achieving thousands of launches per second…
75,000,000,000 Streaming Inserts/Second Using Hierarchical Hypersparse GraphBLAS Matrices
Jeremy Kepner, Tim Davis, Chansup Byun +16
The SuiteSparse GraphBLAS C-library implements high performance hypersparse matrices with bindings to a variety of languages (Python, Julia, and Matlab/Octave). GraphBLAS provides…