4 papers
Denoising GER: A Noise-Robust Generative Error Correction with LLM for Speech Recognition
Yanyan Liu, Minqiang Xu, Yihao Chen +4
In recent years, large language models (LLM) have made significant progress in the task of generation error correction (GER) for automatic speech recognition (ASR) post-processing.…
Wav2DF-TSL: Two-stage Learning with Efficient Pre-training and Hierarchical Experts Fusion for Robust Audio Deepfake Detection
Yunqi Hao, Yihao Chen, Minqiang Xu +5
In recent years, self-supervised learning (SSL) models have made significant progress in audio deepfake detection (ADD) tasks. However, existing SSL models mainly rely on large-sca…
Enhancing Self-Supervised Speaker Verification Using Similarity-Connected Graphs and GCN
Zhaorui Sun, Yihao Chen, Jialong Wang +4
With the continuous development of speech recognition technology, speaker verification (SV) has become an important method for identity authentication. Traditional SV methods rely…
USTC-KXDIGIT System Description for ASVspoof5 Challenge
Yihao Chen, Haochen Wu, Nan Jiang +13
This paper describes the USTC-KXDIGIT system submitted to the ASVspoof5 Challenge for Track 1 (speech deepfake detection) and Track 2 (spoofing-robust automatic speaker verificatio…