activity
20182022
most citedFrom Speaker Verification to Multispeaker Speech Synthesis, Deep Transfer with Feedback Constraint

5 citations · 14 across the 7 of their papers we have counts for

collaborators

10 papers

eess.AS20221 cited

Waveform Boundary Detection for Partially Spoofed Audio

Zexin Cai, Weiqing Wang, Ming Li

The present paper proposes a waveform boundary detection system for audio spoofing attacks containing partially manipulated segments. Partially spoofed/fake audio, where part of th…

eess.AS2022

Invertible Voice Conversion

Zexin Cai, Ming Li

In this paper, we propose an invertible deep learning framework called INVVC for voice conversion. It is designed against the possible threats that inherently come along with voice…

eess.AS2021

The DKU System Description for The Interspeech 2021 Auto-KWS Challenge

Yechen Wang, Yan Jia, Murong Ma +2

This paper introduces the system submitted by the DKU-SMIIP team for the Auto-KWS 2021 Challenge. Our implementation consists of a two-stage keyword spotting system based on query-…

cs.LG2020

Training Wake Word Detection with Synthesized Speech Data on Confusion Words

Yan Jia, Zexin Cai, Murong Ma +4

Confusing-words are commonly encountered in real-life keyword spotting applications, which causes severe degradation of performance due to complex spoken terms and various kinds of…

eess.AS20205 cited

Cross-lingual Multispeaker Text-to-Speech under Limited-Data Scenario

Zexin Cai, Yaogen Yang, Ming Li

Modeling voices for multiple speakers and multiple languages in one text-to-speech system has been a challenge for a long time. This paper presents an extension on Tacotron2 to ach…

eess.AS20205 cited

From Speaker Verification to Multispeaker Speech Synthesis, Deep Transfer with Feedback Constraint

Zexin Cai, Chuxiong Zhang, Ming Li

High-fidelity speech can be synthesized by end-to-end text-to-speech models in recent years. However, accessing and controlling speech attributes such as speaker identity, prosody,…