3 papers
eess.AS2023
RVAE-EM: Generative speech dereverberation based on recurrent variational auto-encoder and convolutive transfer function
Pengyu Wang, Xiaofei Li
In indoor scenes, reverberation is a crucial factor in degrading the perceived quality and intelligibility of speech. In this work, we propose a generative dereverberation method.…
eess.AS2023
Frame-wise streaming end-to-end speaker diarization with non-autoregressive self-attention-based attractors
Di Liang, Nian Shao, Xiaofei Li
This work proposes a frame-wise online/streaming end-to-end neural diarization (FS-EEND) method in a frame-in-frame-out fashion. To frame-wisely detect a flexible number of speaker…
eess.AS2023
FN-SSL: Full-Band and Narrow-Band Fusion for Sound Source Localization
Yabo Wang, Bing Yang, Xiaofei Li
Extracting direct-path spatial features is critical for sound source localization in adverse acoustic environments. This paper proposes a full-band and narrow-band fusion network f…