5 papers · 1 filter
Easy Acceleration with Distributed Arrays
Jeremy Kepner, Chansup Byun, LaToya Anderson +20
High level programming languages and GPU accelerators are powerful enablers for a wide range of applications. Achieving scalable vertical (within a compute node), horizontal (acros…
GPU Sharing with Triples Mode
Chansup Byun, Albert Reuther, LaToya Anderson +19
There is a tremendous amount of interest in AI/ML technologies due to the proliferation of generative AI applications such as ChatGPT. This trend has significantly increased demand…
Hypersparse Traffic Matrices from Suricata Network Flows using GraphBLAS
Michael Houle, Michael Jones, Dan Wallmeyer +8
Hypersparse traffic matrices constructed from network packet source and destination addresses is a powerful tool for gaining insights into network traffic. SuiteSparse: GraphBLAS,…
HPC with Enhanced User Separation
Andrew Prout, Albert Reuther, Michael Houle +19
HPC systems used for research run a wide variety of software and workflows. This software is often written or modified by users to meet the needs of their research projects, and ra…
LLload: Simplifying Real-Time Job Monitoring for HPC Users
Chansup Byun, Julia Mullen, Albert Reuther +16
One of the more complex tasks for researchers using HPC systems is performance monitoring and tuning of their applications. Developing a practice of continuous performance improvem…