7 citations · 11 across the 6 of their papers we have counts for
3 papers · 1 filter
Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition
David M. Chan, Shalini Ghosh, Ariya Rastrow +1
Despite improvements to the generalization performance of automated speech recognition (ASR) models, specializing ASR models for downstream tasks remains a challenging task, primar…
Content-Context Factorized Representations for Automated Speech Recognition
David M. Chan, Shalini Ghosh
Deep neural networks have largely demonstrated their ability to perform automated speech recognition (ASR) by extracting meaningful features from input audio frames. Such features,…
Multi-Modal Pre-Training for Automated Speech Recognition
David M. Chan, Shalini Ghosh, Debmalya Chakrabarty +1
Traditionally, research in automated speech recognition has focused on local-first encoding of audio representations to predict the spoken phonemes in an utterance. Unfortunately,…