85 citations · 369 across the 37 of their papers we have counts for
19 papers · 1 filter
Audio Deepfake Detection: A Survey
Jiangyan Yi, Chenglong Wang, Jianhua Tao +3
Audio deepfake detection is an emerging active topic. A growing number of literatures have aimed to study deepfake detection algorithms and achieved effective performance, the prob…
Do You Remember? Overcoming Catastrophic Forgetting for Fake Audio Detection
Xiaohui Zhang, Jiangyan Yi, Jianhua Tao +2
Current fake audio detection algorithms have achieved promising performances on most datasets. However, their performance may be significantly degraded when dealing with audio of a…
Spatial Reconstructed Local Attention Res2Net with F0 Subband for Fake Speech Detection
Cunhang Fan, Jun Xue, Jianhua Tao +4
The rhythm of bonafide speech is often difficult to replicate, which causes that the fundamental frequency (F0) of synthetic speech is significantly different from that of real spe…
TST: Time-Sparse Transducer for Automatic Speech Recognition
Xiaohui Zhang, Mangui Liang, Zhengkun Tian +2
End-to-end model, especially Recurrent Neural Network Transducer (RNN-T), has achieved great success in speech recognition. However, transducer requires a great memory footprint an…
Boosting Fast and High-Quality Speech Synthesis with Linear Diffusion
Haogeng Liu, Tao Wang, Jie Cao +2
Denoising Diffusion Probabilistic Models have shown extraordinary ability on various generative tasks. However, their slow inference speed renders them impractical in speech synthe…
Low-rank Adaptation Method for Wav2vec2-based Fake Audio Detection
Chenglong Wang, Jiangyan Yi, Xiaohui Zhang +3
Self-supervised speech models are a rapidly developing research topic in fake audio detection. Many pre-trained models can serve as feature extractors, learning richer and higher-l…