80 citations · 242 across the 20 of their papers we have counts for
17 papers · 1 filter
Neural Sound Field Decomposition with Super-resolution of Sound Direction
Qiuqiang Kong, Shilei Liu, Junjie Shi +5
Sound field decomposition predicts waveforms in arbitrary directions using signals from a limited number of microphones as inputs. Sound field decomposition is fundamental to downs…
MIMO Self-attentive RNN Beamformer for Multi-speaker Speech Separation
Xiyun Li, Yong Xu, Meng Yu +4
Recently, our proposed recurrent neural network (RNN) based all deep learning minimum variance distortionless response (ADL-MVDR) beamformer method yielded superior performance ove…
Generalized Spatio-Temporal RNN Beamformer for Target Speech Separation
Yong Xu, Zhuohuang Zhang, Meng Yu +2
Although the conventional mask-based minimum variance distortionless response (MVDR) could reduce the non-linear distortion, the residual noise level of the MVDR separated speech i…
Sound Event Detection of Weakly Labelled Data with CNN-Transformer and Automatic Threshold Optimization
Qiuqiang Kong, Yong Xu, Wenwu Wang +1
Sound event detection (SED) is a task to detect sound events in an audio recording. One challenge of the SED task is that many datasets such as the Detection and Classification of…
End-to-End Multi-Channel Speech Separation
Rongzhi Gu, Jian Wu, Shi-Xiong Zhang +6
The end-to-end approach for single-channel speech separation has been studied recently and shown promising results. This paper extended the previous approach and proposed a new end…
A comprehensive study of speech separation: spectrogram vs waveform separation
Fahimeh Bahmaninezhad, Jian Wu, Rongzhi Gu +4
Speech separation has been studied widely for single-channel close-talk microphone recordings over the past few years; developed solutions are mostly in frequency-domain. Recently,…