most citedInter-APU Communication on AMD MI300A Systems via Infinity Fabric: a Deep Dive

1 citations · 1 across the 4 of their papers we have counts for

collaborators

5 papers

cs.DC2025

Dissecting CPU-GPU Unified Physical Memory on AMD MI300A APUs

Jacob Wahlgren, Gabin Schieffer, Ruimin Shi +4

Discrete GPUs are a cornerstone of HPC and data center systems, requiring management of separate CPU and GPU memory spaces. Unified Virtual Memory (UVM) has been proposed to ease t…

cs.DC20251 cited

Inter-APU Communication on AMD MI300A Systems via Infinity Fabric: a Deep Dive

Gabin Schieffer, Jacob Wahlgren, Ruimin Shi +4

The ever-increasing compute performance of GPU accelerators drives up the need for efficient data movements within HPC applications to sustain performance. Proposed as a solution t…

cs.DC2025

ARM SVE Unleashed: Performance and Insights Across HPC Applications on Nvidia Grace

Ruimin Shi, Gabin Schieffer, Maya Gokhale +3

Vector architectures are essential for boosting computing throughput. ARM provides SVE as the next-generation length-agnostic vector extension beyond traditional fixed-length SIMD.…

cs.DC2024

Disaggregated Memory with SmartNIC Offloading: a Case Study on Graph Processing

Jacob Wahlgren, Gabin Schieffer, Maya Gokhale +2

Disaggregated memory breaks the boundary of monolithic servers to enable memory provisioning on demand. Using network-attached memory to provide memory expansion for memory-intensi…

cs.DC2024

Multi-level Memory-Centric Profiling on ARM Processors with ARM SPE

Samuel Miksits, Ruimin Shi, Maya Gokhale +3

High-end ARM processors are emerging in data centers and HPC systems, posing as a strong contender to x86 machines. Memory-centric profiling is an important approach for dissecting…