activity
20162023
most citedThe DiffuseStyleGesture+ entry to the GENEA Challenge 2023

31 citations · 83 across the 16 of their papers we have counts for

collaborators

22 papers

cs.HC202317 cited

UnifiedGesture: A Unified Gesture Synthesis Model for Multiple Skeletons

Sicheng Yang, Zilin Wang, Zhiyong Wu +8

The automatic co-speech gesture generation draws much attention in computer animation. Previous works designed network structures on individual datasets, which resulted in a lack o…

cs.SD2023

Enhancing the vocal range of single-speaker singing voice synthesis with melody-unsupervised pre-training

Shaohuan Zhou, Xu Li, Zhiyong Wu +2

The single-speaker singing voice synthesis (SVS) usually underperforms at pitch values that are out of the singer's vocal range or associated with limited training samples. Based o…

cs.SD2023

Towards Improving the Expressiveness of Singing Voice Synthesis with BERT Derived Semantic Information

Shaohuan Zhou, Shun Lei, Weiya You +5

This paper presents an end-to-end high-quality singing voice synthesis (SVS) system that uses bidirectional encoder representation from Transformers (BERT) derived semantic embeddi…

cs.SD2023

Towards Spontaneous Style Modeling with Semi-supervised Pre-training for Conversational Text-to-Speech Synthesis

Weiqin Li, Shun Lei, Qiaochu Huang +4

The spontaneous behavior that often occurs in conversations makes speech more human-like compared to reading-style. However, synthesizing spontaneous-style speech is challenging du…

cs.SD2023

Improving Mandarin Prosodic Structure Prediction with Multi-level Contextual Information

Jie Chen, Changhe Song, Deyi Tuo +4

For text-to-speech (TTS) synthesis, prosodic structure prediction (PSP) plays an important role in producing natural and intelligible speech. Although inter-utterance linguistic in…

cs.SD202312 cited

LightGrad: Lightweight Diffusion Probabilistic Model for Text-to-Speech

Jie Chen, Xingchen Song, Zhendong Peng +3

Recent advances in neural text-to-speech (TTS) models bring thousands of TTS applications into daily life, where models are deployed in cloud to provide services for customs. Among…