1 citations · 2 across the 7 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
eess.AS2024
LS-EEND: Long-Form Streaming End-to-End Neural Diarization with Online Attractor Extraction
Di Liang, Xiaofei Li
This work proposes a frame-wise online/streaming end-to-end neural diarization (EEND) method, which detects speaker activities in a frame-in-frame-out fashion. The proposed model m…
cs.SD2024★ 1 cited
Multichannel Long-Term Streaming Neural Speech Enhancement for Static and Moving Speakers
Changsheng Quan, Xiaofei Li
In this work, we extend our previously proposed offline SpatialNet for long-term streaming multichannel speech enhancement in both static and moving speaker scenarios. SpatialNet e…
eess.AS2024
Mel-FullSubNet: Mel-Spectrogram Enhancement for Improving Both Speech Quality and ASR
Rui Zhou, Xian Li, Ying Fang +1
In this work, we propose Mel-FullSubNet, a single-channel Mel-spectrogram denoising and dereverberation network for improving both speech quality and automatic speech recognition (…