collaborators

5 papers

cs.SD2025

Continual Audio Deepfake Detection via Universal Adversarial Perturbation

Wangjie Li, Lin Li, Qingyang Hong

The rapid advancement of speech synthesis and voice conversion technologies has raised significant security concerns in multimedia forensics. Although current detection models demo…

cs.SD2025

SuPseudo: A Pseudo-supervised Learning Method for Neural Speech Enhancement in Far-field Speech Recognition

Longjie Luo, Lin Li, Qingyang Hong

Due to the lack of target speech annotations in real-recorded far-field conversational datasets, speech enhancement (SE) models are typically trained on simulated data. However, th…

cs.SD2025

Pseudo Labels-based Neural Speech Enhancement for the AVSR Task in the MISP-Meeting Challenge

Longjie Luo, Shenghui Lu, Lin Li +1

This paper presents our system for the MISP-Meeting Challenge Track 2. The primary difficulty lies in the dataset, which contains strong background noise, reverberation, overlappin…

cs.SD2025

Cross-attention and Self-attention for Audio-visual Speaker Diarization in MISP-Meeting Challenge

Zhaoyang Li, Haodong Zhou, Longjie Luo +4

This paper presents the system developed for Task 1 of the Multi-modal Information-based Speech Processing (MISP) 2025 Challenge. We introduce CASA-Net, an embedding fusion method…

cs.SD2025

Speaker Diarization with Overlapping Community Detection Using Graph Attention Networks and Label Propagation Algorithm

Zhaoyang Li, Jie Wang, XiaoXiao Li +4

In speaker diarization, traditional clustering-based methods remain widely used in real-world applications. However, these methods struggle with the complex distribution of speaker…