6 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.DC2023
MCR-DL: Mix-and-Match Communication Runtime for Deep Learning
Quentin Anthony, Ammar Ahmad Awan, Jeff Rasley +5
In recent years, the training requirements of many state-of-the-art Deep Learning (DL) models have scaled beyond the compute and memory capabilities of a single processor, and nece…
cs.PF2023★ 2 cited
Performance Characterization of using Quantization for DNN Inference on Edge Devices: Extended Version
Hyunho Ahn, Tian Chen, Nawras Alnaasan +5
Quantization is a popular technique used in Deep Neural Networks (DNN) inference to reduce the size of models and improve the overall numerical performance by exploiting native har…
cs.DC2016★ 6 cited
System-level Scalable Checkpoint-Restart for Petascale Computing
Jiajun Cao, Kapil Arya, Rohan Garg +5
Fault tolerance for the upcoming exascale generation has long been an area of active research. One of the components of a fault tolerance strategy is checkpointing. Petascale-level…