3 papers
cs.SD2026
Disentangled Global-Local Feature Learning with E-Branchformer for Audio Deepfake Detection
Phuong Tuan Dat, Ho Bao Thu, Nguyen Tran Trung +2
The rapid advancement of voice synthesis technologies such as text-to-speech and voice conversion poses significant threats to speech-based authentication systems, necessitating ro…
cs.SD2024
VoxVietnam: a Large-Scale Multi-Genre Dataset for Vietnamese Speaker Recognition
Hoang Long Vu, Phuong Tuan Dat, Pham Thao Nhi +2
Recent research in speaker recognition aims to address vulnerabilities due to variations between enrolment and test utterances, particularly in the multi-genre phenomenon where the…
cs.CL2023
The 2022 NIST Language Recognition Evaluation
Yooyoung Lee, Craig Greenberg, Eliot Godard +5
In 2022, the U.S. National Institute of Standards and Technology (NIST) conducted the latest Language Recognition Evaluation (LRE) in an ongoing series administered by NIST since 1…