95 citations
- Shenzhen Technology UniversityCN9 papers
- Chang Gung Memorial HospitalTW6 papers
- Beijing University of Posts and TelecommunicationsCN5 papers
- University of Science and Technology of ChinaCN5 papers
- National Institutes of Health Clinical CenterUS4 papers
- Beijing Institute of TechnologyCN3 papers
- Johns Hopkins UniversityUS3 papers
- Alibaba Group (China)CN2 papers
- Association for Computing MachineryUS2 papers
- China Medical UniversityTW2 papers
- First Affiliated Hospital Zhejiang UniversityCN2 papers
- Huazhong University of Science and TechnologyCN2 papers
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024★ 13 cited
Retrieval-Augmented Audio Deepfake Detection
Zuheng Kang, Yayun He, Botao Zhao +4
With recent advances in speech synthesis including text-to-speech (TTS) and voice conversion (VC) systems enabling the generation of ultra-realistic audio deepfakes, there is growi…
cs.SD2023★ 15 cited
PMVC: Data Augmentation-Based Prosody Modeling for Expressive Voice Conversion
Yimin Deng, Huaizhen Tang, Xulong Zhang +3
Voice conversion as the style transfer task applied to speech, refers to converting one person's speech into a new speech that sounds like another person's. Up to now, there has be…
cs.SD2022★ 19 cited
TGAVC: Improving Autoencoder Voice Conversion with Text-Guided and Adversarial Training
Huaizhen Tang, Xulong Zhang, Jianzong Wang +4
Non-parallel many-to-many voice conversion remains an interesting but challenging speech processing task. Recently, AutoVC, a conditional autoencoder based method, achieved excelle…