collaborators
Showing cs.DCShow all

6 papers · 1 filter

cs.DC2025

Easy Acceleration with Distributed Arrays

Jeremy Kepner, Chansup Byun, LaToya Anderson +20

High level programming languages and GPU accelerators are powerful enablers for a wide range of applications. Achieving scalable vertical (within a compute node), horizontal (acros…

cs.DC2024

GPU Sharing with Triples Mode

Chansup Byun, Albert Reuther, LaToya Anderson +19

There is a tremendous amount of interest in AI/ML technologies due to the proliferation of generative AI applications such as ChatGPT. This trend has significantly increased demand…

cs.DC2024

Supercomputer 3D Digital Twin for User Focused Real-Time Monitoring

William Bergeron, Matthew Hubbell, Daniel Mojica +17

Real-time supercomputing performance analysis is a critical aspect of evaluating and optimizing computational systems in a dynamic user environment. The operation of supercomputers…

cs.DC2024

Hypersparse Traffic Matrices from Suricata Network Flows using GraphBLAS

Michael Houle, Michael Jones, Dan Wallmeyer +8

Hypersparse traffic matrices constructed from network packet source and destination addresses is a powerful tool for gaining insights into network traffic. SuiteSparse: GraphBLAS,…

cs.DC2024

HPC with Enhanced User Separation

Andrew Prout, Albert Reuther, Michael Houle +19

HPC systems used for research run a wide variety of software and workflows. This software is often written or modified by users to meet the needs of their research projects, and ra…

cs.DC2024

LLload: Simplifying Real-Time Job Monitoring for HPC Users

Chansup Byun, Julia Mullen, Albert Reuther +16

One of the more complex tasks for researchers using HPC systems is performance monitoring and tuning of their applications. Developing a practice of continuous performance improvem…