14 citations · 66 across the 24 of their papers we have counts for
4 papers · 1 filter
SLD-L2S: Hierarchical Subspace Latent Diffusion for High-Fidelity Lip to Speech Synthesis
Yifan Liang, Andong Li, Kang Yang +5
Although lip-to-speech synthesis (L2S) has achieved significant progress in recent years, current state-of-the-art methods typically rely on intermediate representations such as me…
Low-latency Monaural Speech Enhancement with Deep Filter-bank Equalizer
Chengshi Zheng, Wenzhe Liu, Andong Li +2
It is highly desirable that speech enhancement algorithms can achieve good performance while keeping low latency for many applications, such as digital hearing aids, acoustically t…
A deep complex multi-frame filtering network for stereophonic acoustic echo cancellation
Linjuan Cheng, Chengshi Zheng, Andong Li +3
In hands-free communication system, the coupling between loudspeaker and microphone generates echo signal, which can severely influence the quality of communication. Meanwhile, var…
Incorporating Multi-Target in Multi-Stage Speech Enhancement Model for Better Generalization
Lu Zhang, Mingjiang Wang, Andong Li +2
Recent single-channel speech enhancement methods based on deep neural networks (DNNs) have achieved remarkable results, but there are still generalization problems in real scenes.…