1 citations · 1 across the 7 of their papers we have counts for
7 papers
The First Voice Timbre Attribute Detection Challenge
Liping Chen, Jinghao He, Zhengyan Sheng +2
The first voice timbre attribute detection challenge is featured in a special session at NCMMSC 2025. It focuses on the explainability of voice timbre and compares the intensity of…
Introducing voice timbre attribute detection
Jinghao He, Zhengyan Sheng, Liping Chen +2
This paper focuses on explaining the timbre conveyed by speech signals and introduces a task termed voice timbre attribute detection (vTAD). In this task, voice timbre is explained…
The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan
Zhengyan Sheng, Jinghao He, Liping Chen +2
Voice timbre refers to the unique quality or character of a person's voice that distinguishes it from others as perceived by human hearing. The Voice Timbre Attribute Detection (Vt…
Incremental Disentanglement for Environment-Aware Zero-Shot Text-to-Speech Synthesis
Ye-Xin Lu, Hui-Peng Du, Zheng-Yan Sheng +2
This paper proposes an Incremental Disentanglement-based Environment-Aware zero-shot text-to-speech (TTS) method, dubbed IDEA-TTS, that can synthesize speech for unseen speakers wh…
Multi-Stage Speech Bandwidth Extension with Flexible Sampling Rate Control
Ye-Xin Lu, Yang Ai, Zheng-Yan Sheng +1
The majority of existing speech bandwidth extension (BWE) methods operate under the constraint of fixed source and target sampling rates, which limits their flexibility in practica…
Voice Attribute Editing with Text Prompt
Zhengyan Sheng, Yang Ai, Li-Juan Liu +2
Despite recent advancements in speech generation with text prompt providing control over speech style, voice attributes in synthesized speech remain elusive and challenging to cont…