3 papers
cs.SD2026
DisSR: Disentangling Speech Representation for Degradation-Prior Guided Cross-Domain Speech Restoration
Ziqi Liang, Zhijun Jia, Chang Liu +3
Previous speech restoration (SR) primarily focuses on single-task speech restoration (SSR), which cannot address general speech restoration problems. Training specific SSR models f…
cs.SD2025
CycleFlow: Leveraging Cycle Consistency in Flow Matching for Speaker Style Adaptation
Ziqi Liang, Xulong Zhang, Chang Liu +3
Voice Conversion (VC) aims to convert the style of a source speaker, such as timbre and pitch, to the style of any target speaker while preserving the linguistic content. However,…
cs.CL2024
AlignCap: Aligning Speech Emotion Captioning to Human Preferences
Ziqi Liang, Haoxiang Shi, Hanhui Chen
Speech Emotion Captioning (SEC) has gradually become an active research task. The emotional content conveyed through human speech are often complex, and classifying them into fixed…