2 papers
eess.AS2025
Multi-Scale Accent Modeling and Disentangling for Multi-Speaker Multi-Accent Text-to-Speech Synthesis
Xuehao Zhou, Mingyang Zhang, Yi Zhou +2
Generating speech across different accents while preserving speaker identity is crucial for various real-world applications. However, accurately and independently modeling both spe…
eess.AS2024
RefXVC: Cross-Lingual Voice Conversion with Enhanced Reference Leveraging
Mingyang Zhang, Yi Zhou, Yi Ren +3
This paper proposes RefXVC, a method for cross-lingual voice conversion (XVC) that leverages reference information to improve conversion performance. Previous XVC works generally t…