activity
20182022
most citedA comparison of end-to-end models for long-form speech recognition

13 citations · 22 across the 9 of their papers we have counts for

collaborators
Showing eess.ASShow all

13 papers · 1 filter

eess.AS2022

A Universally-Deployable ASR Frontend for Joint Acoustic Echo Cancellation, Speech Enhancement, and Voice Separation

Tom O'Malley, Arun Narayanan, Quan Wang

Recent work has shown that it is possible to train a single model to perform joint acoustic echo cancellation (AEC), speech enhancement, and voice separation, thereby serving as a…

eess.AS2022

Streaming Noise Context Aware Enhancement For Automatic Speech Recognition in Multi-Talker Environments

Joe Caroselli, Arun Narayanan, Yiteng Huang

One of the most challenging scenarios for smart speakers is multi-talker, when target speech from the desired speaker is mixed with interfering speech from one or more speakers. A…

eess.AS2022

A Conformer-based Waveform-domain Neural Acoustic Echo Canceller Optimized for ASR Accuracy

Sankaran Panchapagesan, Arun Narayanan, Turaj Zakizadeh Shabestary +5

Acoustic Echo Cancellation (AEC) is essential for accurate recognition of queries spoken to a smart speaker that is playing out audio. Previous work has shown that a neural AEC mod…

eess.AS20223 cited

Mask scalar prediction for improving robust automatic speech recognition

Arun Narayanan, James Walker, Sankaran Panchapagesan +2

Using neural network based acoustic frontends for improving robustness of streaming automatic speech recognition (ASR) systems is challenging because of the causality constraints a…

eess.AS20211 cited

Cross-attention conformer for context modeling in speech enhancement for ASR

Arun Narayanan, Chung-Cheng Chiu, Tom O'Malley +2

This work introduces \emph{cross-attention conformer}, an attention-based architecture for context modeling in speech enhancement. Given that the context information can often be s…

eess.AS2021

Personalized Keyphrase Detection using Speaker and Environment Information

Rajeev Rikhye, Quan Wang, Qiao Liang +6

In this paper, we introduce a streaming keyphrase detection system that can be easily customized to accurately detect any phrase composed of words from a large vocabulary. The syst…