6 citations · 19 across the 7 of their papers we have counts for
4 papers · 1 filter
E-Branchformer: Branchformer with Enhanced merging for speech recognition
Kwangyoun Kim, Felix Wu, Yifan Peng +4
Conformer, combining convolution and self-attention sequentially to capture both local and global information, has shown remarkable performance and is currently regarded as the sta…
SRU++: Pioneering Fast Recurrence with Attention for Speech Recognition
Jing Pan, Tao Lei, Kwangyoun Kim +2
The Transformer architecture has been well adopted as a dominant architecture in most sequence transduction tasks including automatic speech recognition (ASR), since its attention…
ASAPP-ASR: Multistream CNN and Self-Attentive SRU for SOTA Speech Recognition
Jing Pan, Joshua Shapiro, Jeremy Wohlwend +3
In this paper we present state-of-the-art (SOTA) performance on the LibriSpeech corpus with two novel neural network architectures, a multistream CNN for acoustic modeling and a se…
Multistream CNN for Robust Acoustic Modeling
Kyu J. Han, Jing Pan, Venkata Krishna Naveen Tadala +2
This paper proposes multistream CNN, a novel neural network architecture for robust acoustic modeling in speech recognition tasks. The proposed architecture processes input speech…