collaborators

6 papers

eess.AS2025

Listen to Extract: Onset-Prompted Target Speaker Extraction

Pengjie Shen, Kangrui Chen, Shulin He +5

We propose listen to extract (LExt), a highly-effective while extremely-simple algorithm for monaural target speaker extraction (TSE). Given an enrollment utterance of a target spe…

eess.AS2025

ARiSE: Auto-Regressive Multi-Channel Speech Enhancement

Pengjie Shen, Xueliang Zhang, Zhong-Qiu Wang

We propose ARiSE, an auto-regressive algorithm for multi-channel speech enhancement. ARiSE improves existing deep neural network (DNN) based frame-online multi-channel speech enhan…

cs.SD2025

Multi-Channel Acoustic Echo Cancellation Based on Direction-of-Arrival Estimation

Fei Zhao, Xueliang Zhang, Zhong-Qiu Wang

Acoustic echo cancellation (AEC) is an important speech signal processing technology that can remove echoes from microphone signals to enable natural-sounding full-duplex speech co…

cs.SD2025

Room Impulse Response as a Prompt for Acoustic Echo Cancellation

Fei Zhao, Shulin He, Xueliang Zhang

Data-driven acoustic echo cancellation (AEC) methods, predominantly trained on synthetic or constrained real-world datasets, encounter performance declines in unseen echo scenarios…

cs.SD2024

Robust Target Speaker Direction of Arrival Estimation

Zixuan Li, Shulin He, Xueliang Zhang

In multi-speaker environments the direction of arrival (DOA) of a target speaker is key for improving speech clarity and extracting target speaker's voice. However, traditional DOA…

cs.SD2024

Attention-Enhanced Short-Time Wiener Solution for Acoustic Echo Cancellation

Fei Zhao, Xueliang Zhang

Acoustic Echo Cancellation (AEC) is an essential speech signal processing technology that removes echoes from microphone inputs to facilitate natural-sounding full-duplex communica…