74 citations · 123 across the 2 of their papers we have counts for
2 papers
cs.LG2017★ 74 cited
AdaComp : Adaptive Residual Gradient Compression for Data-Parallel Distributed Training
Chia-Yu Chen, Jungwook Choi, Daniel Brand +3
Highly distributed training of Deep Neural Networks (DNNs) on future compute platforms (offering 100 of TeraOps/s of computational capacity) is expected to be severely communicatio…
cs.LG2017★ 49 cited
MEC: Memory-efficient Convolution for Deep Neural Network
Minsik Cho, Daniel Brand
Convolution is a critical component in modern deep neural networks, thus several algorithms for convolution have been developed. Direct convolution is simple but suffers from poor…