Showing cs.DCShow all
3 papers · 1 filter
cs.DC2025
Characterizing Production GPU Workloads using System-wide Telemetry Data
Onur Cankur, Brian Austin, Dhruva Kulkarni +1
GPGPU-accelerated clusters and supercomputers are central to modern high-performance computing (HPC). Over the past decade, these systems continue to expand, and GPUs now expose a…
cs.DC2024
Automated Programmatic Performance Analysis of Parallel Programs
Onur Cankur, Aditya Tomar, Daniel Nichols +3
Developing efficient parallel applications is critical to advancing scientific development but requires significant performance analysis and optimization. Performance analysis tool…
cs.DC2023
Pipit: Scripting the analysis of parallel execution traces
Abhinav Bhatele, Rakrish Dhakal, Alexander Movsesyan +2
Performance analysis is a critical step in the oft-repeated, iterative process of performance tuning of parallel programs. Per-process, per-thread traces (detailed logs of events w…