collaborators

8 papers

cs.DC2026

Stencil Computations on Cerebras Wafer-Scale Engine

Elia Belli, Daniele De Sensi

Stencil computations are a fundamental kernel in scientific computing, critical for simulations in domains such as fluid dynamics and climate modeling. However, these computations…

cs.DC2026

Stencil Computations on Tenstorrent Wormhole

Lorenzo Piarulli, Daniele De Sensi

As investment in AI-focused accelerators grows and their deployment in supercomputing facilities expands, understanding whether these architectures can efficiently support traditio…

cs.DC2026

The Landscape of GPU-Centric Communication

Didem Unat, Ilyas Turimbetov, Mohammed Kefah Taha Issa +4

In recent years, GPUs have become the preferred accelerators for HPC and ML applications due to their parallelism and fast memory bandwidth. While GPUs boost computation, inter-GPU…

cs.DC2026

Characterizing the Impact of Congestion in Modern HPC Interconnects

Lorenzo Piarulli, Marco Faltelli, Dirk Pleiter +7

High-performance computing (HPC) systems increasingly support both scalable AI training and large-scale simulation workloads. Both typically rely heavily on collective communicatio…

cs.DC2026

PICO: Performance Insights for Collective Operations

Saverio Pasqualoni, Tommaso Bonato, Lorenzo Piarulli +3

Collective operations are cornerstones of both HPC applications and large-scale AI training and inference, yet benchmarking them in a systematic and reproducible way remains diffic…

cs.NI2026

SMaRTT: Sender-based Marked Rapidly-adapting Trimmed & Timed Transport

Tommaso Bonato, Abdul Kabbani, Ahmad Ghalayini +10

With the rapid growth of artificial intelligence (AI) workloads in datacenters, the Ultra Ethernet Consortium (UEC) has defined a new high-performance transport layer to deliver th…