activity
20232025
most citedFace-Driven Zero-Shot Voice Conversion with Memory-based Face-Voice Alignment

1 citations · 1 across the 7 of their papers we have counts for

collaborators

7 papers

cs.SD2025

The First Voice Timbre Attribute Detection Challenge

Liping Chen, Jinghao He, Zhengyan Sheng +2

The first voice timbre attribute detection challenge is featured in a special session at NCMMSC 2025. It focuses on the explainability of voice timbre and compares the intensity of…

cs.SD2025

Introducing voice timbre attribute detection

Jinghao He, Zhengyan Sheng, Liping Chen +2

This paper focuses on explaining the timbre conveyed by speech signals and introduces a task termed voice timbre attribute detection (vTAD). In this task, voice timbre is explained…

cs.SD2025

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan

Zhengyan Sheng, Jinghao He, Liping Chen +2

Voice timbre refers to the unique quality or character of a person's voice that distinguishes it from others as perceived by human hearing. The Voice Timbre Attribute Detection (Vt…

eess.AS2024

Incremental Disentanglement for Environment-Aware Zero-Shot Text-to-Speech Synthesis

Ye-Xin Lu, Hui-Peng Du, Zheng-Yan Sheng +2

This paper proposes an Incremental Disentanglement-based Environment-Aware zero-shot text-to-speech (TTS) method, dubbed IDEA-TTS, that can synthesize speech for unseen speakers wh…

eess.AS2024

Multi-Stage Speech Bandwidth Extension with Flexible Sampling Rate Control

Ye-Xin Lu, Yang Ai, Zheng-Yan Sheng +1

The majority of existing speech bandwidth extension (BWE) methods operate under the constraint of fixed source and target sampling rates, which limits their flexibility in practica…

cs.SD2024

Voice Attribute Editing with Text Prompt

Zhengyan Sheng, Yang Ai, Li-Juan Liu +2

Despite recent advancements in speech generation with text prompt providing control over speech style, voice attributes in synthesized speech remain elusive and challenging to cont…