2 papers
eess.AS2025
DialoSpeech: Dual-Speaker Dialogue Generation with LLM and Flow Matching
Hanke Xie, Dake Guo, Chengyou Wang +8
Recent advances in text-to-speech (TTS) synthesis, particularly those leveraging large language models (LLMs), have significantly improved expressiveness and naturalness. However,…
cs.SD2020
The SLT 2021 children speech recognition challenge: Open datasets, rules and baselines
Fan Yu, Zhuoyuan Yao, Xiong Wang +6
Automatic speech recognition (ASR) has been significantly advanced with the use of deep learning and big data. However improving robustness, including achieving equally good perfor…