3 papers
cs.CL2023
Retrieve and Copy: Scaling ASR Personalization to Large Catalogs
Sai Muralidhar Jayanthi, Devang Kulshreshtha, Saket Dingliwal +2
Personalization of automatic speech recognition (ASR) models is a widely studied topic because of its many practical applications. Most recently, attention-based contextual biasing…
cs.SD2023
Generalized zero-shot audio-to-intent classification
Veera Raghavendra Elluru, Devang Kulshreshtha, Rohit Paturi +2
Spoken language understanding systems using audio-only data are gaining popularity, yet their ability to handle unseen intents remains limited. In this study, we propose a generali…
eess.AS2023
Dynamic Chunk Convolution for Unified Streaming and Non-Streaming Conformer ASR
Xilai Li, Goeric Huybrechts, Srikanth Ronanki +2
Recently, there has been an increasing interest in unifying streaming and non-streaming speech recognition models to reduce development, training and deployment cost. The best-know…