5 citations · 6 across the 5 of their papers we have counts for
5 papers
UniSyn: An End-to-End Unified Model for Text-to-Speech and Singing Voice Synthesis
Yi Lei, Shan Yang, Xinsheng Wang +4
Text-to-speech (TTS) and singing voice synthesis (SVS) aim at generating high-quality speaking and singing voice according to textual input and music scores, respectively. Unifying…
Expressive-VC: Highly Expressive Voice Conversion with Attention Fusion of Bottleneck and Perturbation Features
Ziqian Ning, Qicong Xie, Pengcheng Zhu +5
Voice conversion for highly expressive speech is challenging. Current approaches struggle with the balancing between speaker similarity, intelligibility and expressiveness. To addr…
Distinguishable Speaker Anonymization based on Formant and Fundamental Frequency Scaling
Jixun Yao, Qing Wang, Yi Lei +4
Speech data on the Internet are proliferating exponentially because of the emergence of social media, and the sharing of such personal data raises obvious security and privacy conc…
Preserving background sound in noise-robust voice conversion via multi-task learning
Jixun Yao, Yi Lei, Qing Wang +6
Background sound is an informative form of art that is helpful in providing a more immersive experience in real-application voice conversion (VC) scenarios. However, prior research…
NWPU-ASLP System for the VoicePrivacy 2022 Challenge
Jixun Yao, Qing Wang, Li Zhang +3
This paper presents the NWPU-ASLP speaker anonymization system for VoicePrivacy 2022 Challenge. Our submission does not involve additional Automatic Speaker Verification (ASV) mode…