activity
20182022
most citedCUP: Cluster Pruning for Compressing Deep Neural Networks

9 citations · 12 across the 3 of their papers we have counts for

collaborators
Showing cs.DCShow all

5 papers · 1 filter

cs.DC2022

"Smarter" NICs for faster molecular dynamics: a case study

Sara Karamati, Clayton Hughes, K. Scott Hemmert +6

This work evaluates the benefits of using a "smart" network interface card (SmartNIC) as a compute accelerator for the example of the MiniMD molecular dynamics proxy application. T…

cs.DC20193 cited

Load-Balanced Sparse MTTKRP on GPUs

Israt Nisa, Jiajia Li, Aravind Sukumaran-Rajam +2

Sparse matricized tensor times Khatri-Rao product (MTTKRP) is one of the most computationally expensive kernels in sparse tensor computations. This work focuses on optimizing the M…

cs.DC2018

Programming Strategies for Irregular Algorithms on the Emu Chick

Eric Hein, Srinivas Eswar, Abdurrahman Yaşar +7

The Emu Chick prototype implements migratory memory-side processing in a novel hardware system. Rather than transferring large amounts of data across the system interconnect, the E…

cs.DC2018

A Microbenchmark Characterization of the Emu Chick

Jeffrey S. Young, Eric Hein, Srinivas Eswar +5

The Emu Chick is a prototype system designed around the concept of migratory memory-side processing. Rather than transferring large amounts of data across power-hungry, high-latenc…

cs.DC2018

Accurate, Fast and Scalable Kernel Ridge Regression on Parallel and Distributed Systems

Yang You, James Demmel, Cho-Jui Hsieh +1

We propose two new methods to address the weak scaling problems of KRR: the Balanced KRR (BKRR) and K-means KRR (KKRR). These methods consider alternative ways to partition the inp…