collaborators

5 papers

cs.DC2025

Characterizing Adaptive Mesh Refinement on Heterogeneous Platforms with Parthenon-VIBE

Akash Poptani, Alireza Khadem, Scott Mahlke +3

Hero-class HPC simulations rely on Adaptive Mesh Refinement (AMR) to reduce compute and memory demands while maintaining accuracy. This work analyzes the performance of Parthenon,…

cs.AR2025

A Customized Memory-aware Architecture for Biological Sequence Alignment

Nasrin Akbari, Mehdi Modarressi, Alireza Khadem

Sequence alignment is a fundamental process in computational biology which identifies regions of similarity in biological sequences. With the exponential growth in the volume of da…

cs.AR2025

DX100: A Programmable Data Access Accelerator for Indirection

Alireza Khadem, Kamalavasan Kamalakkannan, Zhenyan Zhu +8

Indirect memory accesses frequently appear in applications where memory bandwidth is a critical bottleneck. Prior indirect memory access proposals, such as indirect prefetchers, ru…

cs.AR2025

PIM Is All You Need: A CXL-Enabled GPU-Free System for Large Language Model Inference

Yufeng Gu, Alireza Khadem, Sumanth Umesh +5

Large Language Model (LLM) inference uses an autoregressive manner to generate one token at a time, which exhibits notably lower operational intensity compared to earlier Machine L…

cs.AR2025

Multi-Dimensional Vector ISA Extension for Mobile In-Cache Computing

Alireza Khadem, Daichi Fujiki, Hilbert Chen +4

In-cache computing technology transforms existing caches into long-vector compute units and offers low-cost alternatives to building expensive vector engines for mobile CPUs. Unfor…