9 papers
Communication Offloading on SmartNIC DPUs: A Quantitative Approach
Jacob Wahlgren, Andong Hu, Roger Pearce +2
SmartNIC Data Processing Units (DPUs) offer a promising solution for saving high-end CPU resources by offloading tasks to programmable cores near the network interface. In this wor…
Dissecting CPU-GPU Unified Physical Memory on AMD MI300A APUs
Jacob Wahlgren, Gabin Schieffer, Ruimin Shi +4
Discrete GPUs are a cornerstone of HPC and data center systems, requiring management of separate CPU and GPU memory spaces. Unified Virtual Memory (UVM) has been proposed to ease t…
Inter-APU Communication on AMD MI300A Systems via Infinity Fabric: a Deep Dive
Gabin Schieffer, Jacob Wahlgren, Ruimin Shi +4
The ever-increasing compute performance of GPU accelerators drives up the need for efficient data movements within HPC applications to sustain performance. Proposed as a solution t…
ARC-V: Vertical Resource Adaptivity for HPC Workloads in Containerized Environments
Daniel Medeiros, Jeremy J. Williams, Jacob Wahlgren +2
Existing state-of-the-art vertical autoscalers for containerized environments are traditionally built for cloud applications, which might behave differently than HPC workloads with…
Kub: Enabling Elastic HPC Workloads on Containerized Environments
Daniel Medeiros, Jacob Wahlgren, Gabin Schieffer +1
The conventional model of resource allocation in HPC systems is static. Thus, a job cannot leverage newly available resources in the system or release underutilized resources durin…
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
Daniel Medeiros, Gabin Schieffer, Jacob Wahlgren +1
Complex workflows play a critical role in accelerating scientific discovery. In many scientific domains, efficient workflow management can lead to faster scientific output and broa…