5 citations · 14 across the 7 of their papers we have counts for
10 papers
Waveform Boundary Detection for Partially Spoofed Audio
Zexin Cai, Weiqing Wang, Ming Li
The present paper proposes a waveform boundary detection system for audio spoofing attacks containing partially manipulated segments. Partially spoofed/fake audio, where part of th…
Invertible Voice Conversion
Zexin Cai, Ming Li
In this paper, we propose an invertible deep learning framework called INVVC for voice conversion. It is designed against the possible threats that inherently come along with voice…
The DKU System Description for The Interspeech 2021 Auto-KWS Challenge
Yechen Wang, Yan Jia, Murong Ma +2
This paper introduces the system submitted by the DKU-SMIIP team for the Auto-KWS 2021 Challenge. Our implementation consists of a two-stage keyword spotting system based on query-…
Training Wake Word Detection with Synthesized Speech Data on Confusion Words
Yan Jia, Zexin Cai, Murong Ma +4
Confusing-words are commonly encountered in real-life keyword spotting applications, which causes severe degradation of performance due to complex spoken terms and various kinds of…
Cross-lingual Multispeaker Text-to-Speech under Limited-Data Scenario
Zexin Cai, Yaogen Yang, Ming Li
Modeling voices for multiple speakers and multiple languages in one text-to-speech system has been a challenge for a long time. This paper presents an extension on Tacotron2 to ach…
From Speaker Verification to Multispeaker Speech Synthesis, Deep Transfer with Feedback Constraint
Zexin Cai, Chuxiong Zhang, Ming Li
High-fidelity speech can be synthesized by end-to-end text-to-speech models in recent years. However, accessing and controlling speech attributes such as speaker identity, prosody,…