6 papers
Listen to Extract: Onset-Prompted Target Speaker Extraction
Pengjie Shen, Kangrui Chen, Shulin He +5
We propose listen to extract (LExt), a highly-effective while extremely-simple algorithm for monaural target speaker extraction (TSE). Given an enrollment utterance of a target spe…
ARiSE: Auto-Regressive Multi-Channel Speech Enhancement
Pengjie Shen, Xueliang Zhang, Zhong-Qiu Wang
We propose ARiSE, an auto-regressive algorithm for multi-channel speech enhancement. ARiSE improves existing deep neural network (DNN) based frame-online multi-channel speech enhan…
Multi-Channel Acoustic Echo Cancellation Based on Direction-of-Arrival Estimation
Fei Zhao, Xueliang Zhang, Zhong-Qiu Wang
Acoustic echo cancellation (AEC) is an important speech signal processing technology that can remove echoes from microphone signals to enable natural-sounding full-duplex speech co…
Room Impulse Response as a Prompt for Acoustic Echo Cancellation
Fei Zhao, Shulin He, Xueliang Zhang
Data-driven acoustic echo cancellation (AEC) methods, predominantly trained on synthetic or constrained real-world datasets, encounter performance declines in unseen echo scenarios…
Robust Target Speaker Direction of Arrival Estimation
Zixuan Li, Shulin He, Xueliang Zhang
In multi-speaker environments the direction of arrival (DOA) of a target speaker is key for improving speech clarity and extracting target speaker's voice. However, traditional DOA…
Attention-Enhanced Short-Time Wiener Solution for Acoustic Echo Cancellation
Fei Zhao, Xueliang Zhang
Acoustic Echo Cancellation (AEC) is an essential speech signal processing technology that removes echoes from microphone inputs to facilitate natural-sounding full-duplex communica…