papers

Publications (9)

cs.DC2025

Easy Acceleration with Distributed Arrays

Jeremy Kepner, Chansup Byun, LaToya Anderson +20

High level programming languages and GPU accelerators are powerful enablers for a wide range of applications. Achieving scalable vertical (within a compute node), horizontal (acros…

cs.DC2024

HPC with Enhanced User Separation

Andrew Prout, Albert Reuther, Michael Houle +19

HPC systems used for research run a wide variety of software and workflows. This software is often written or modified by users to meet the needs of their research projects, and ra…

cs.PF2024

LLload: An Easy-to-Use HPC Utilization Tool

Chansup Byun, Albert Reuther, Julie Mullen +19

The increasing use and cost of high performance computing (HPC) requires new easy-to-use tools to enable HPC users and HPC systems engineers to transparently understand the utiliza…

cs.DC2024

Supercomputer 3D Digital Twin for User Focused Real-Time Monitoring

William Bergeron, Matthew Hubbell, Daniel Mojica +17

Real-time supercomputing performance analysis is a critical aspect of evaluating and optimizing computational systems in a dynamic user environment. The operation of supercomputers…

cs.DC2024

GPU Sharing with Triples Mode

Chansup Byun, Albert Reuther, LaToya Anderson +19

There is a tremendous amount of interest in AI/ML technologies due to the proliferation of generative AI applications such as ChatGPT. This trend has significantly increased demand…

cs.DC2024

LLload: Simplifying Real-Time Job Monitoring for HPC Users

Chansup Byun, Julia Mullen, Albert Reuther +16

One of the more complex tasks for researchers using HPC systems is performance monitoring and tuning of their applications. Developing a practice of continuous performance improvem…