14 citations · 66 across the 24 of their papers we have counts for
24 papers · 1 filter
Scalable Neural Vocoder from Range-Null Space Decomposition
Andong Li, Tong Lei, Zhihang Sun +4
Although deep neural networks have facilitated significant progress of neural vocoders in recent years, they usually suffer from intrinsic challenges like opaque modeling, inflexib…
GOMPSNR: Reflourish the Signal-to-Noise Ratio Metric for Audio Generation Tasks
Lingling Dai, Andong Li, Cheng Chi +3
In the field of audio generation, signal-to-noise ratio (SNR) has long served as an objective metric for evaluating audio quality. Nevertheless, recent studies have shown that SNR…
NaturalL2S: End-to-End High-quality Multispeaker Lip-to-Speech Synthesis with Differential Digital Signal Processing
Yifan Liang, Fangkun Liu, Andong Li +2
Recent advancements in visual speech recognition (VSR) have promoted progress in lip-to-speech synthesis, where pre-trained VSR models enhance the intelligibility of synthesized sp…
Neural Vocoders as Speech Enhancers
Andong Li, Zhihang Sun, Fengyuan Hao +2
Speech enhancement (SE) and neural vocoding are traditionally viewed as separate tasks. In this work, we observe them under a common thread: the rank behavior of these processes. T…
BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement
Cunhang Fan, Enrui Liu, Andong Li +5
Although the complex spectrum-based speech enhancement(SE) methods have achieved significant performance, coupling amplitude and phase can lead to a compensation effect, where ampl…
SMRU: Split-and-Merge Recurrent-based UNet for Acoustic Echo Cancellation and Noise Suppression
Zhihang Sun, Andong Li, Rilin Chen +4
The proliferation of deep neural networks has spawned the rapid development of acoustic echo cancellation and noise suppression, and plenty of prior arts have been proposed, which…