26 citations · 42 across the 6 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
Clustering and Mining Accented Speech for Inclusive and Fair Speech Recognition
Jaeyoung Kim, Han Lu, Soheil Khorram +3
Modern automatic speech recognition (ASR) systems are typically trained on more than tens of thousands hours of speech data, which is one of the main factors for their great succes…
cs.SD2020
Transformer Transducer: One Model Unifying Streaming and Non-streaming Speech Recognition
Anshuman Tripathi, Jaeyoung Kim, Qian Zhang +2
In this paper we present a Transformer-Transducer model architecture and a training technique to unify streaming and non-streaming speech recognition models into one model. The mod…