1.1k citations · 1.2k across the 2 of their papers we have counts for
2 papers
cs.LG2017★ 74 cited
AdaComp : Adaptive Residual Gradient Compression for Data-Parallel Distributed Training
Chia-Yu Chen, Jungwook Choi, Daniel Brand +3
Highly distributed training of Deep Neural Networks (DNNs) on future compute platforms (offering 100 of TeraOps/s of computational capacity) is expected to be severely communicatio…
cs.LG2015★ 1.1k cited
Deep Learning with Limited Numerical Precision
Suyog Gupta, Ankur Agrawal, Kailash Gopalakrishnan +1
Training of large-scale deep neural networks is often constrained by the available computational resources. We study the effect of limited precision data representation and computa…