activity
20172022
most citedMask scalar prediction for improving robust automatic speech recognition

3 citations · 3 across the 2 of their papers we have counts for

collaborators

5 papers

eess.AS2022

A Conformer-based Waveform-domain Neural Acoustic Echo Canceller Optimized for ASR Accuracy

Sankaran Panchapagesan, Arun Narayanan, Turaj Zakizadeh Shabestary +5

Acoustic Echo Cancellation (AEC) is essential for accurate recognition of queries spoken to a smart speaker that is playing out audio. Previous work has shown that a neural AEC mod…

eess.AS20223 cited

Mask scalar prediction for improving robust automatic speech recognition

Arun Narayanan, James Walker, Sankaran Panchapagesan +2

Using neural network based acoustic frontends for improving robustness of streaming automatic speech recognition (ASR) systems is challenging because of the causality constraints a…

eess.AS2020

Efficient Knowledge Distillation for RNN-Transducer Models

Sankaran Panchapagesan, Daniel S. Park, Chung-Cheng Chiu +3

Knowledge Distillation is an effective method of transferring knowledge from a large model to a smaller model. Distillation can be viewed as a type of model compression, and has pl…

cs.CL2018

Data Augmentation for Robust Keyword Spotting under Playback Interference

Anirudh Raju, Sankaran Panchapagesan, Xing Liu +2

Accurate on-device keyword spotting (KWS) with low false accept and false reject rate is crucial to customer experience for far-field voice control of conversational agents. It is…

cs.CL2017

Max-Pooling Loss Training of Long Short-Term Memory Networks for Small-Footprint Keyword Spotting

Ming Sun, Anirudh Raju, George Tucker +6

We propose a max-pooling based loss function for training Long Short-Term Memory (LSTM) networks for small-footprint keyword spotting (KWS), with low CPU, memory, and latency requi…