1 citations · 1 across the 2 of their papers we have counts for
4 papers
Scaling All-to-all Operations Across Emerging Many-Core Supercomputers
Shannon Kinkead, Jackson Wesley, Whit Schonbein +3
Performant all-to-all collective operations in MPI are critical to fast Fourier transforms, transposition, and machine learning applications. There are many existing implementation…
Persistent and Partitioned MPI for Stencil Communication
Gerald Collom, Jason Burmark, Olga Pearce +1
Many parallel applications rely on iterative stencil operations, whose performance are dominated by communication costs at large scales. Several MPI optimizations, such as persiste…
MPI Advance : Open-Source Message Passing Optimizations
Amanda Bienz, Derek Schafer, Anthony Skjellum
The large variety of production implementations of the message passing interface (MPI) each provide unique and varying underlying algorithms. Each emerging supercomputer supports o…
Optimizing Irregular Communication with Neighborhood Collectives and Locality-Aware Parallelism
Gerald Collom, Rui Peng Li, Amanda Bienz
Irregular communication often limits both the performance and scalability of parallel applications. Typically, applications individually implement irregular messages using point-to…