2 papers
eess.AS2025
PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data
Songjun Cao, Qinghua Wu, Jie Chen +2
As parallel training data is scarce for one-shot voice conversion (VC) tasks, waveform reconstruction is typically performed by various VC systems. A typical one-shot VC system com…
cs.SD2025
DiffCSS: Diverse and Expressive Conversational Speech Synthesis with Diffusion Models
Weihao wu, Zhiwei Lin, Yixuan Zhou +6
Conversational speech synthesis (CSS) aims to synthesize both contextually appropriate and expressive speech, and considerable efforts have been made to enhance the understanding o…