2 citations · 4 across the 2 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2025
Text-Queried Audio Source Separation via Hierarchical Modeling
Xinlei Yin, Xiulian Peng, Xue Jiang +2
Target audio source separation with natural language queries presents a promising paradigm for extracting arbitrary audio events through arbitrary text descriptions. Existing metho…
cs.SD2024★ 2 cited
Convert and Speak: Zero-shot Accent Conversion with Minimum Supervision
Zhijun Jia, Huaying Xue, Xiulian Peng +1
Low resource of parallel data is the key challenge of accent conversion(AC) problem in which both the pronunciation units and prosody pattern need to be converted. We propose a two…
cs.SD2022★ 2 cited
End-to-End Neural Speech Coding for Real-Time Communications
Xue Jiang, Xiulian Peng, Chengyu Zheng +3
Deep-learning based methods have shown their advantages in audio coding over traditional ones but limited attention has been paid on real-time communications (RTC). This paper prop…