5 citations · 10 across the 18 of their papers we have counts for
4 papers · 1 filter
Retrieve and Copy: Scaling ASR Personalization to Large Catalogs
Sai Muralidhar Jayanthi, Devang Kulshreshtha, Saket Dingliwal +2
Personalization of automatic speech recognition (ASR) models is a widely studied topic because of its many practical applications. Most recently, attention-based contextual biasing…
Generalized zero-shot audio-to-intent classification
Veera Raghavendra Elluru, Devang Kulshreshtha, Rohit Paturi +2
Spoken language understanding systems using audio-only data are gaining popularity, yet their ability to handle unseen intents remains limited. In this study, we propose a generali…
DCTX-Conformer: Dynamic context carry-over for low latency unified streaming and non-streaming Conformer ASR
Goeric Huybrechts, Srikanth Ronanki, Xilai Li +3
Conformer-based end-to-end models have become ubiquitous these days and are commonly used in both streaming and non-streaming automatic speech recognition (ASR). Techniques like du…
Dynamic Chunk Convolution for Unified Streaming and Non-Streaming Conformer ASR
Xilai Li, Goeric Huybrechts, Srikanth Ronanki +2
Recently, there has been an increasing interest in unifying streaming and non-streaming speech recognition models to reduce development, training and deployment cost. The best-know…