13 citations · 13 across the 2 of their papers we have counts for
2 papers
cs.CL2021
4-bit Quantization of LSTM-based Speech Recognition Models
Andrea Fasoli, Chia-Yu Chen, Mauricio Serrano +9
We investigate the impact of aggressive low-precision representations of weights and activations in two families of large LSTM-based architectures for Automatic Speech Recognition…
cs.LG2021★ 13 cited
ScaleCom: Scalable Sparsified Gradient Compression for Communication-Efficient Distributed Training
Chia-Yu Chen, Jiamin Ni, Songtao Lu +8
Large-scale distributed training of Deep Neural Networks (DNNs) on state-of-the-art platforms is expected to be severely communication constrained. To overcome this limitation, num…