2 papers
cs.CL2020
Improved Robustness to Disfluencies in RNN-Transducer Based Speech Recognition
Valentin Mendelev, Tina Raissi, Guglielmo Camporese +1
Automatic Speech Recognition (ASR) based on Recurrent Neural Network Transducers (RNN-T) is gaining interest in the speech community. We investigate data selection and preparation…
eess.AS2020
Bootstrap an end-to-end ASR system by multilingual training, transfer learning, text-to-text mapping and synthetic audio
Manuel Giollo, Deniz Gunceler, Yulan Liu +1
Bootstrapping speech recognition on limited data resources has been an area of active research for long. The recent transition to all-neural models and end-to-end (E2E) training br…