14 citations · 15 across the 2 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2020★ 1 cited
Audio-Visual Decision Fusion for WFST-based and seq2seq Models
Rohith Aralikatti, Sharad Roy, Abhinav Thanda +4
Under noisy conditions, speech recognition systems suffer from high Word Error Rates (WER). In such cases, information from the visual modality comprising the speaker lip movements…
eess.AS2018
Global SNR Estimation of Speech Signals using Entropy and Uncertainty Estimates from Dropout Networks
Rohith Aralikatti, Dilip Margam, Tanay Sharma +2
This paper demonstrates two novel methods to estimate the global SNR of speech signals. In both methods, Deep Neural Network-Hidden Markov Model (DNN-HMM) acoustic model used in sp…