2 citations · 3 across the 5 of their papers we have counts for
6 papers
Fast and Scalable Sparse Triangular Solver for Multi-GPU Based HPC Architectures
Chenhao Xie, Jieyang Chen, Jesun S Firoz +5
Designing efficient and scalable sparse linear algebra kernels on modern multi-GPU based HPC systems is a daunting task due to significant irregular memory references and workload…
Leaky Buddies: Cross-Component Covert Channels on Integrated CPU-GPU Systems
Sankha Baran Dutta, Hoda Naghibijouybari, Nael Abu-Ghazaleh +2
Graphics Processing Units (GPUs) are a ubiquitous component across the range of today's computing platforms, from phones and tablets, through personal computers, to high-end server…
ARENA: Asynchronous Reconfigurable Accelerator Ring to Enable Data-Centric Parallel Computing
Cheng Tan, Chenhao Xie, Tong Geng +4
The next generation HPC and data centers are likely to be reconfigurable and data-centric due to the trend of hardware specialization and the emergence of data-driven applications.…
A Parallel Sparse Tensor Benchmark Suite on CPUs and GPUs
Jiajia Li, Mahesh Lakshminarasimhan, Xiaolong Wu +3
Tensor computations present significant performance challenges that impact a wide spectrum of applications ranging from machine learning, healthcare analytics, social network analy…
Evaluating Modern GPU Interconnect: PCIe, NVLink, NV-SLI, NVSwitch and GPUDirect
Ang Li, Shuaiwen Leon Song, Jieyang Chen +4
High performance multi-GPU computing becomes an inevitable trend due to the ever-increasing demand on computation capability in emerging domains such as deep learning, big data and…
PASTA: A Parallel Sparse Tensor Algorithm Benchmark Suite
Jiajia Li, Yuchen Ma, Xiaolong Wu +2
Tensor methods have gained increasingly attention from various applications, including machine learning, quantum chemistry, healthcare analytics, social network analysis, data mini…