2 papers
cs.PL2026
Linear Layouts: Robust Code Generation of Efficient Tensor Computation Using
Keren Zhou, Mario Lezcano, Adam Goucher +8
Efficient tensor computation is a cornerstone of modern deep learning (DL) workloads, yet existing approaches struggle to achieve flexible and performant design and implementation…
cs.DC2026
PASTA: A Modular Program Analysis Tool Framework for Accelerators
Mao Lin, Hyeran Jeon, Keren Zhou
The increasing complexity and diversity of hardware accelerators in modern computing systems demand flexible, low-overhead program analysis tools. We present PASTA, a low-overhead…