38 citations · 137 across the 21 of their papers we have counts for
4 papers · 1 filter
ASR Error Correction and Domain Adaptation Using Machine Translation
Anirudh Mani, Shruti Palaskar, Nimshi Venkat Meripo +2
Off-the-shelf pre-trained Automatic Speech Recognition (ASR) systems are an increasingly viable service for companies of any size building speech-based products. While these ASR sy…
Cross-Attention End-to-End ASR for Two-Party Conversations
Suyoun Kim, Siddharth Dalmia, Florian Metze
We present an end-to-end speech recognition model that learns interaction between two speakers based on the turn-changing information. Unlike conventional speech recognition models…
Acoustic-to-Word Recognition with Sequence-to-Sequence Models
Shruti Palaskar, Florian Metze
Acoustic-to-Word recognition provides a straightforward solution to end-to-end speech recognition without needing external decoding, language model re-scoring or lexicon. While cha…
End-to-End Multimodal Speech Recognition
Shruti Palaskar, Ramon Sanabria, Florian Metze
Transcription or sub-titling of open-domain videos is still a challenging domain for Automatic Speech Recognition (ASR) due to the data's challenging acoustics, variable signal pro…