1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.SD2026
AugCodec: A Low-Bitrate Disentangled Neural Speech Codec via Data Augmentation
Dongmei Wang, Xiaohang Sun, Yang Liu +10
We propose AugCodec, a low-bitrate disentangled neural speech codec that leverages data augmentation to decompose speech into three distinct components: semantic, speaker, and pros…
cs.SD2022
Egocentric Audio-Visual Noise Suppression
Roshan Sharma, Weipeng He, Ju Lin +3
This paper studies audio-visual noise suppression for egocentric videos -- where the speaker is not captured in the video. Instead, potential noise sources are visible on screen wi…
cs.SD2022★ 1 cited
LPCSE: Neural Speech Enhancement through Linear Predictive Coding
Yang Liu, Na Tang, Xiaoli Chu +2
The increasingly stringent requirement on quality-of-experience in 5G/B5G communication systems has led to the emerging neural speech enhancement techniques, which however have bee…