3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2025
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
Jun-Hyeok Cha, Seung-Bin Kim, Hyung-Seok Oh +1
Recently, there has been a growing demand for conversational speech synthesis (CSS) that generates more natural speech by considering the conversational context. To address this, w…
cs.SD2023★ 3 cited
HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
Sang-Hoon Lee, Ha-Yeong Choi, Seung-Bin Kim +1
Large language models (LLM)-based speech synthesis has been widely adopted in zero-shot speech synthesis. However, they require a large-scale data and possess the same limitations…