collaborators

7 papers

cs.SD2024

RaD-Net 2: A causal two-stage repairing and denoising speech enhancement network with knowledge distillation and complex axial self-attention

Mingshuai Liu, Zhuangqi Chen, Xiaopeng Yan +5

In real-time speech communication systems, speech signals are often degraded by multiple distortions. Recently, a two-stage Repair-and-Denoising network (RaD-Net) was proposed with…

eess.AS2024

BS-PLCNet 2: Two-stage Band-split Packet Loss Concealment Network with Intra-model Knowledge Distillation

Zihan Zhang, Xianjun Xia, Chuanzeng Huang +2

Audio packet loss is an inevitable problem in real-time speech communication. A band-split packet loss concealment network (BS-PLCNet) targeting full-band signals was recently prop…

cs.SD2024

RaD-Net: A Repairing and Denoising Network for Speech Signal Improvement

Mingshuai Liu, Zhuangqi Chen, Xiaopeng Yan +5

This paper introduces our repairing and denoising network (RaD-Net) for the ICASSP 2024 Speech Signal Improvement (SSI) Challenge. We extend our previous framework based on a two-s…

eess.AS2024

BS-PLCNet: Band-split Packet Loss Concealment Network with Multi-task Learning Framework and Multi-discriminators

Zihan Zhang, Jiayao Sun, Xianjun Xia +3

Packet loss is a common and unavoidable problem in voice over internet phone (VoIP) systems. To deal with the problem, we propose a band-split packet loss concealment network (BS-P…

eess.AS2023

An Exploration of Task-decoupling on Two-stage Neural Post Filter for Real-time Personalized Acoustic Echo Cancellation

Zihan Zhang, Jiayao Sun, Xianjun Xia +4

Deep learning based techniques have been popularly adopted in acoustic echo cancellation (AEC). Utilization of speaker representation has extended the frontier of AEC, thus attract…

eess.AS2023

Harmonic enhancement using learnable comb filter for light-weight full-band speech enhancement model

Xiaohuai Le, Tong Lei, Li Chen +9

With fewer feature dimensions, filter banks are often used in light-weight full-band speech enhancement models. In order to further enhance the coarse speech in the sub-band domain…