activity
20182020
most citedStreaming End-to-End Bilingual ASR Systems with Joint Language Identification

9 citations · 22 across the 4 of their papers we have counts for

collaborators

7 papers

eess.AS20204 cited

Streaming ResLSTM with Causal Mean Aggregation for Device-Directed Utterance Detection

Xiaosu Tong, Che-Wei Huang, Sri Harish Mallidi +5

In this paper, we propose a streaming model to distinguish voice queries intended for a smart-home device from background speech. The proposed model consists of multiple CNN layers…

eess.AS20209 cited

Streaming End-to-End Bilingual ASR Systems with Joint Language Identification

Surabhi Punjabi, Harish Arsikere, Zeynab Raeesy +11

Multilingual ASR technology simplifies model training and deployment, but its accuracy is known to depend on the availability of language information at runtime. Since language ide…

cs.CL2020

Neural Machine Translation for Multilingual Grapheme-to-Phoneme Conversion

Alex Sokolov, Tracy Rohlin, Ariya Rastrow

Grapheme-to-phoneme (G2P) models are a key component in Automatic Speech Recognition (ASR) systems, such as the ASR system in Alexa, as they are used to generate pronunciations for…

eess.AS20209 cited

Streaming Language Identification using Combination of Acoustic Representations and ASR Hypotheses

Chander Chandak, Zeynab Raeesy, Ariya Rastrow +5

This paper presents our modeling and architecture approaches for building a highly accurate low-latency language identification system to support multilingual spoken queries for vo…

eess.AS2019

Audio-attention discriminative language model for ASR rescoring

Ankur Gandhe, Ariya Rastrow

End-to-end approaches for automatic speech recognition (ASR) benefit from directly modeling the probability of the word sequence given the input audio stream in a single neural net…

cs.CL2019

Scalable Multi Corpora Neural Language Models for ASR

Anirudh Raju, Denis Filimonov, Gautam Tiwari +2

Neural language models (NLM) have been shown to outperform conventional n-gram language models by a substantial margin in Automatic Speech Recognition (ASR) and other tasks. There…