collaborators
Showing cs.DCShow all

6 papers · 1 filter

cs.DC2026

Implementing True MPI Sessions and Evaluating MPI Initialization Scalability

Hui Zhou, Kenneth Raffenetti, Yanfei Guo +2

Sessions is one of the major features introduced in the MPI-4 standard. It offers an alternative to the traditional world communicator model by allowing applications to construct c…

cs.DC20253 cited

FIRST: Federated Inference Resource Scheduling Toolkit for Scientific AI Model Access

Aditya Tanikanti, Benoit Côté, Yanfei Guo +9

We present the Federated Inference Resource Scheduling Toolkit (FIRST), a framework enabling Inference-as-a-Service across distributed High-Performance Computing (HPC) clusters. FI…

cs.DC2025

Examining MPI and its Extensions for Asynchronous Multithreaded Communication

Jiakun Yan, Marc Snir, Yanfei Guo

The increasing complexity of HPC architectures and the growing adoption of irregular scientific algorithms demand efficient support for asynchronous, multithreaded communication. T…

cs.DC2025

ZCCL: Significantly Improving Collective Communication With Error-Bounded Lossy Compression

Jiajun Huang, Sheng Di, Xiaodong Yu +12

With the ever-increasing computing power of supercomputers and the growing scale of scientific applications, the efficiency of MPI collective communication turns out to be a critic…

cs.DC2024

MPI Progress For All

Hui Zhou, Robert Latham, Ken Raffenetti +2

The progression of communication in the Message Passing Interface (MPI) is not well defined, yet it is critical for application performance, particularly in achieving effective com…

cs.DC2024

Designing and Prototyping Extensions to MPI in MPICH

Hui Zhou, Ken Raffenetti, Yanfei Guo +3

As HPC system architectures and the applications running on them continue to evolve, the MPI standard itself must evolve. The trend in current and future HPC systems toward powerfu…