15 citations · 36 across the 7 of their papers we have counts for
4 papers · 1 filter
Waveform Boundary Detection for Partially Spoofed Audio
Zexin Cai, Weiqing Wang, Ming Li
The present paper proposes a waveform boundary detection system for audio spoofing attacks containing partially manipulated segments. Partially spoofed/fake audio, where part of th…
The DKU-DukeECE Diarization System for the VoxCeleb Speaker Recognition Challenge 2022
Weiqing Wang, Xiaoyi Qin, Ming Cheng +3
This paper discribes the DKU-DukeECE submission to the 4th track of the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22). Our system contains a fused voice activity detectio…
Invertible Voice Conversion
Zexin Cai, Ming Li
In this paper, we propose an invertible deep learning framework called INVVC for voice conversion. It is designed against the possible threats that inherently come along with voice…
Cross-lingual Multispeaker Text-to-Speech under Limited-Data Scenario
Zexin Cai, Yaogen Yang, Ming Li
Modeling voices for multiple speakers and multiple languages in one text-to-speech system has been a challenge for a long time. This paper presents an extension on Tacotron2 to ach…