3 papers
cs.SD2026
DisSR: Disentangling Speech Representation for Degradation-Prior Guided Cross-Domain Speech Restoration
Ziqi Liang, Zhijun Jia, Chang Liu +3
Previous speech restoration (SR) primarily focuses on single-task speech restoration (SSR), which cannot address general speech restoration problems. Training specific SSR models f…
eess.AS2025
Group Relative Policy Optimization for Text-to-Speech with Large Language Models
Chang Liu, Ya-Jun Hu, Ying-Ying Gao +2
This paper proposes a GRPO-based approach to enhance the performance of large language model (LLM)-based text-to-speech (TTS) models by deriving rewards from an off-the-shelf autom…
cs.SD2025
CycleFlow: Leveraging Cycle Consistency in Flow Matching for Speaker Style Adaptation
Ziqi Liang, Xulong Zhang, Chang Liu +3
Voice Conversion (VC) aims to convert the style of a source speaker, such as timbre and pitch, to the style of any target speaker while preserving the linguistic content. However,…