2 papers
eess.AS2025
Deep Dubbing: End-to-End Auto-Audiobook System with Text-to-Timbre and Context-Aware Instruct-TTS
Ziqi Dai, Yiting Chen, Jiacheng Xu +8
The pipeline for multi-participant audiobook production primarily consists of three stages: script analysis, character voice timbre selection, and speech synthesis. Among these, sc…
cs.SD2025
Open-Source Full-Duplex Conversational Datasets for Natural and Interactive Speech Synthesis
Zhitong Zhou, Qingqing Zhang, Lei Luo +2
Full-duplex, spontaneous conversational data are essential for enhancing the naturalness and interactivity of synthesized speech in conversational TTS systems. We present two open-…