7 citations · 18 across the 7 of their papers we have counts for
7 papers
Sparsely Overlapped Speech Training in the Time Domain: Joint Learning of Target Speech Separation and Personal VAD Benefits
Qingjian Lin, Lin Yang, Xuyang Wang +3
Target speech separation is the process of filtering a certain speaker's voice out of speech mixtures according to the additional speaker identity information provided. Recent work…
TransfoRNN: Capturing the Sequential Information in Self-Attention Representations for Language Modeling
Tze Yuang Chong, Xuyang Wang, Lin Yang +1
In this paper, we describe the use of recurrent neural networks to capture sequential information from the self-attention representations to improve the Transformers. Although self…
The 2020 Personalized Voice Trigger Challenge: Open Database, Evaluation Metrics and the Baseline Systems
Yan Jia, Xingming Wang, Xiaoyi Qin +4
The 2020 Personalized Voice Trigger Challenge (PVTC2020) addresses two different research problems a unified setup: joint wake-up word detection with speaker verification on close-…
Exploring Voice Conversion based Data Augmentation in Text-Dependent Speaker Verification
Xiaoyi Qin, Yaogen Yang, Lin Yang +3
In this paper, we focus on improving the performance of the text-dependent speaker verification system in the scenario of limited training data. The speaker verification system dee…
Training Wake Word Detection with Synthesized Speech Data on Confusion Words
Yan Jia, Zexin Cai, Murong Ma +4
Confusing-words are commonly encountered in real-life keyword spotting applications, which causes severe degradation of performance due to complex spoken terms and various kinds of…
Mask Detection and Breath Monitoring from Speech: on Data Augmentation, Feature Representation and Modeling
Haiwei Wu, Lin Zhang, Lin Yang +4
This paper introduces our approaches for the Mask and Breathing Sub-Challenge in the Interspeech COMPARE Challenge 2020. For the mask detection task, we train deep convolutional ne…