113 citations · 348 across the 32 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2021
Zero-Shot Text-to-Speech for Text-Based Insertion in Audio Narration
Chuanxin Tang, Chong Luo, Zhiyuan Zhao +3
Given a piece of speech and its transcript text, text-based speech editing aims to generate speech that can be seamlessly inserted into the given speech by editing the transcript.…
cs.SD2021★ 8 cited
General-Purpose Speech Representation Learning through a Self-Supervised Multi-Granularity Framework
Yucheng Zhao, Dacheng Yin, Chong Luo +4
This paper presents a self-supervised learning framework, named MGF, for general-purpose speech representation learning. In the design of MGF, speech hierarchy is taken into consid…
cs.SD2019★ 26 cited
PHASEN: A Phase-and-Harmonics-Aware Speech Enhancement Network
Dacheng Yin, Chong Luo, Zhiwei Xiong +1
Time-frequency (T-F) domain masking is a mainstream approach for single-channel speech enhancement. Recently, focuses have been put to phase prediction in addition to amplitude pre…