16 citations · 88 across the 30 of their papers we have counts for
6 papers · 1 filter
SLD-L2S: Hierarchical Subspace Latent Diffusion for High-Fidelity Lip to Speech Synthesis
Yifan Liang, Andong Li, Kang Yang +5
Although lip-to-speech synthesis (L2S) has achieved significant progress in recent years, current state-of-the-art methods typically rely on intermediate representations such as me…
Rethinking the joint estimation of magnitude and phase for time-frequency domain neural vocoders
Lingling Dai, Andong Li, Tong Lei +3
Time-frequency (T-F) domain-based neural vocoders have shown promising results in synthesizing high-fidelity audio. Nevertheless, it remains unclear on the mechanism of effectively…
Low-latency Monaural Speech Enhancement with Deep Filter-bank Equalizer
Chengshi Zheng, Wenzhe Liu, Andong Li +2
It is highly desirable that speech enhancement algorithms can achieve good performance while keeping low latency for many applications, such as digital hearing aids, acoustically t…
A deep complex multi-frame filtering network for stereophonic acoustic echo cancellation
Linjuan Cheng, Chengshi Zheng, Andong Li +3
In hands-free communication system, the coupling between loudspeaker and microphone generates echo signal, which can severely influence the quality of communication. Meanwhile, var…
A Robust Maximum Likelihood Distortionless Response Beamformer based on a Complex Generalized Gaussian Distribution
Weixin Meng, Chengshi Zheng, Xiaodong Li
For multichannel speech enhancement, this letter derives a robust maximum likelihood distortionless response beamformer by modeling speech sparse priors with a complex generalized…
Distributed Node-Specific Block-Diagonal LCMV Beamforming in Wireless Acoustic Sensor Networks
Xinwei Guo, Minmin Yuan, Chengshi Zheng +1
This paper derives the analytical solution of a novel distributed node-specific block-diagonal linearly constrained minimum variance beamformer from the centralized linearly constr…