2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 1 cited
Accelerator-Aware Training for Transducer-Based Speech Recognition
Suhaila M. Shakiah, Rupak Vignesh Swaminathan, Hieu Duy Nguyen +6
Machine learning model weights and activations are represented in full-precision during training. This leads to performance degradation in runtime when deployed on neural network a…
eess.AS2022★ 2 cited
Sub-8-Bit Quantization Aware Training for 8-Bit Neural Network Accelerator with On-Device Speech Recognition
Kai Zhen, Hieu Duy Nguyen, Raviteja Chinta +4
We present a novel sub-8-bit quantization-aware training (S8BQAT) scheme for 8-bit neural network accelerators. Our method is inspired from Lloyd-Max compression theory with practi…