Showing eess.ASShow all
2 papers · 1 filter
eess.AS2025
XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation
Tianlun Zuo, Jingbin Hu, Yuke Li +6
Zero-shot emotion transfer in cross-lingual speech synthesis refers to generating speech in a target language, where the emotion is expressed based on reference speech from a diffe…
eess.AS2022
Preserving background sound in noise-robust voice conversion via multi-task learning
Jixun Yao, Yi Lei, Qing Wang +6
Background sound is an informative form of art that is helpful in providing a more immersive experience in real-application voice conversion (VC) scenarios. However, prior research…