2 papers
cs.DC2026
GICC: A High-Performance Runtime for GPU-Initiated Communication and Coordination in Modern HPC Systems
Baodi Shan, Mauricio Araya-Polo, Barbara Chapman
Distributed GPU applications increasingly rely on kernel-level, cross-node coordination to reduce launch overheads and improve compute-communication overlap, but such support is la…
cs.DC2025
DiOMP-Offloading: Toward Portable Distributed Heterogeneous OpenMP
Baodi Shan, Mauricio Araya-Polo, Barbara Chapman
As core counts and heterogeneity rise in HPC, traditional hybrid programming models face challenges in managing distributed GPU memory and ensuring portability. This paper presents…