Publications (4)
Bootstrap an end-to-end ASR system by multilingual training, transfer learning, text-to-text mapping and synthetic audio
Manuel Giollo, Deniz Gunceler, Yulan Liu +1
Bootstrapping speech recognition on limited data resources has been an area of active research for long. The recent transition to all-neural models and end-to-end (E2E) training br…
Hallucination Benchmark for Speech Foundation Models
Alkis Koudounas, Moreno La Quatra, Manuel Giollo +2
Hallucinations in automatic speech recognition (ASR) systems refer to fluent and coherent transcriptions produced by neural ASR models that are completely unrelated to the underlyi…
An expanded evaluation of protein function prediction methods shows an improvement in accuracy
Yuxiang Jiang, Tal Ronnen Oron, Wyatt T Clark +144
Background: The increasing volume and variety of genotypic and phenotypic data is a major defining characteristic of modern biomedical sciences. At the same time, the limitations i…
Improved Robustness to Disfluencies in RNN-Transducer Based Speech Recognition
Valentin Mendelev, Tina Raissi, Guglielmo Camporese +1
Automatic Speech Recognition (ASR) based on Recurrent Neural Network Transducers (RNN-T) is gaining interest in the speech community. We investigate data selection and preparation…