14 citations · 19 across the 19 of their papers we have counts for
Showing 2023 · eess.ASShow all
2 papers · 2 filters
eess.AS2023
Lookahead When It Matters: Adaptive Non-causal Transformers for Streaming Neural Transducers
Grant P. Strimel, Yi Xie, Brian King +3
Streaming speech recognition architectures are employed for low-latency, real-time applications. Such architectures are often characterized by their causality. Causal architectures…
eess.AS2023★ 14 cited
PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
Rahul Pandey, Roger Ren, Qi Luo +7
End-to-End (E2E) automatic speech recognition (ASR) systems used in voice assistants often have difficulties recognizing infrequent words personalized to the user, such as names an…