9 citations · 10 across the 2 of their papers we have counts for
2 papers
eess.AS2023★ 1 cited
Accented Text-to-Speech Synthesis with Limited Data
Xuehao Zhou, Mingyang Zhang, Yi Zhou +2
This paper presents an accented text-to-speech (TTS) synthesis framework with limited training data. We study two aspects concerning accent rendering: phonetic (phoneme difference)…
cs.SD2023★ 9 cited
AUDIT: Audio Editing by Following Instructions with Latent Diffusion Models
Yuancheng Wang, Zeqian Ju, Xu Tan +4
Audio editing is applicable for various purposes, such as adding background sound effects, replacing a musical instrument, and repairing damaged audio. Recently, some diffusion-bas…