2 citations · 2 across the 3 of their papers we have counts for
3 papers
eess.AS2024
FLY-TTS: Fast, Lightweight and High-Quality End-to-End Text-to-Speech Synthesis
Yinlin Guo, Yening Lv, Jinqiao Dou +2
While recent advances in Text-To-Speech synthesis have yielded remarkable improvements in generating high-quality speech, research on lightweight and fast models is limited. This p…
cs.SD2024
Continuous Target Speech Extraction: Enhancing Personalized Diarization and Extraction on Complex Recordings
He Zhao, Hangting Chen, Jianwei Yu +1
Target speaker extraction (TSE) aims to extract the target speaker's voice from the input mixture. Previous studies have concentrated on high-overlapping scenarios. However, real-w…
eess.AS2024★ 2 cited
Audio Deepfake Detection with Self-Supervised WavLM and Multi-Fusion Attentive Classifier
Yinlin Guo, Haofan Huang, Xi Chen +2
With the rapid development of speech synthesis and voice conversion technologies, Audio Deepfake has become a serious threat to the Automatic Speaker Verification (ASV) system. Num…