3 papers
cs.PF2022
Machine Learning for CUDA+MPI Design Rules
Carl Pearson, Aurya Javeed, Karen Devine
We present a new strategy for automatically exploring the design space of key CUDA+MPI programs and providing design rules that discriminate slow from fast implementations. In such…
cs.DC2021
Parallel Graph Coloring Algorithms for Distributed GPU Environments
Ian Bogle, Erik G Boman, Karen D Devine +2
Graph coloring is often used in parallelizing scientific computations that run in distributed and multi-GPU environments; it identifies sets of independent data that can be updated…
cs.DC2018
Geometric Partitioning and Ordering Strategies for Task Mapping on Parallel Computers
Mehmet Deveci, Karen D. Devine, Kevin Pedretti +3
We present a new method for mapping applications' MPI tasks to cores of a parallel computer such that applications' communication time is reduced. We address the case of sparse nod…