1 citations · 4 across the 25 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2024
Text Prompt is Not Enough: Sound Event Enhanced Prompt Adapter for Target Style Audio Generation
Chenxu Xiong, Ruibo Fu, Shuchen Shi +9
Current mainstream audio generation methods primarily rely on simple text prompts, often failing to capture the nuanced details necessary for multi-style audio generation. To addre…
eess.AS2022
Unsupervised Quantized Prosody Representation for Controllable Speech Synthesis
Yutian Wang, Yuankun Xie, Kun Zhao +2
In this paper, we propose a novel prosody disentangle method for prosodic Text-to-Speech (TTS) model, which introduces the vector quantization (VQ) method to the auxiliary prosody…