activity
20162022
most citedLingvo: a Modular and Scalable Framework for Sequence-to-Sequence Modeling

184 citations · 193 across the 6 of their papers we have counts for

collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS2022

Closing the Gap between Single-User and Multi-User VoiceFilter-Lite

Rajeev Rikhye, Quan Wang, Qiao Liang +2

VoiceFilter-Lite is a speaker-conditioned voice separation model that plays a crucial role in improving speech recognition and speaker verification by suppressing overlapping speec…

eess.AS2021

Multi-user VoiceFilter-Lite via Attentive Speaker Embedding

Rajeev Rikhye, Quan Wang, Qiao Liang +2

In this paper, we propose a solution to allow speaker conditioned speech models, such as VoiceFilter-Lite, to support an arbitrary number of enrolled users in a single pass. This i…

eess.AS2021

Multi-Task Learning for End-to-End ASR Word and Utterance Confidence with Deletion Prediction

David Qiu, Yanzhang He, Qiujia Li +3

Confidence scores are very useful for downstream applications of automatic speech recognition (ASR) systems. Recent works have proposed using neural networks to learn word or utter…

eess.AS2021

Personalized Keyphrase Detection using Speaker and Environment Information

Rajeev Rikhye, Quan Wang, Qiao Liang +6

In this paper, we introduce a streaming keyphrase detection system that can be easily customized to accurately detect any phrase composed of words from a large vocabulary. The syst…

eess.AS20212 cited

Learning Word-Level Confidence For Subword End-to-End ASR

David Qiu, Qiujia Li, Yanzhang He +9

We study the problem of word-level confidence estimation in subword-based end-to-end (E2E) models for automatic speech recognition (ASR). Although prior works have proposed trainin…