8 papers
Stencil Computations on Cerebras Wafer-Scale Engine
Elia Belli, Daniele De Sensi
Stencil computations are a fundamental kernel in scientific computing, critical for simulations in domains such as fluid dynamics and climate modeling. However, these computations…
Stencil Computations on Tenstorrent Wormhole
Lorenzo Piarulli, Daniele De Sensi
As investment in AI-focused accelerators grows and their deployment in supercomputing facilities expands, understanding whether these architectures can efficiently support traditio…
The Landscape of GPU-Centric Communication
Didem Unat, Ilyas Turimbetov, Mohammed Kefah Taha Issa +4
In recent years, GPUs have become the preferred accelerators for HPC and ML applications due to their parallelism and fast memory bandwidth. While GPUs boost computation, inter-GPU…
Characterizing the Impact of Congestion in Modern HPC Interconnects
Lorenzo Piarulli, Marco Faltelli, Dirk Pleiter +7
High-performance computing (HPC) systems increasingly support both scalable AI training and large-scale simulation workloads. Both typically rely heavily on collective communicatio…
PICO: Performance Insights for Collective Operations
Saverio Pasqualoni, Tommaso Bonato, Lorenzo Piarulli +3
Collective operations are cornerstones of both HPC applications and large-scale AI training and inference, yet benchmarking them in a systematic and reproducible way remains diffic…
SMaRTT: Sender-based Marked Rapidly-adapting Trimmed & Timed Transport
Tommaso Bonato, Abdul Kabbani, Ahmad Ghalayini +10
With the rapid growth of artificial intelligence (AI) workloads in datacenters, the Ultra Ethernet Consortium (UEC) has defined a new high-performance transport layer to deliver th…