10 papers
BreathNet: Generalizable Audio Deepfake Detection via Breath-Cue-Guided Feature Refinement
Zhe Ye, Xiangui Kang, Jiayi He +5
As deepfake audio becomes more realistic and diverse, developing generalizable countermeasure systems has become crucial. Existing detection methods primarily depend on XLS-R front…
SLRTP2025 Sign Language Production Challenge: Methodology, Results, and Future Work
Harry Walsh, Ed Fish, Ozge Mercanoglu Sincan +19
Sign Language Production (SLP) is the task of generating sign language video from spoken language inputs. The field has seen a range of innovations over the last few years, with th…
Motion is the Choreographer: Learning Latent Pose Dynamics for Seamless Sign Language Generation
Jiayi He, Xu Wang, Shengeng Tang +3
Sign language video generation requires producing natural signing motions with realistic appearances under precise semantic control, yet faces two critical challenges: excessive si…
ALLM4ADD: Unlocking the Capabilities of Audio Large Language Models for Audio Deepfake Detection
Hao Gu, Jiangyan Yi, Chenglong Wang +6
Audio deepfake detection (ADD) has grown increasingly important due to the rise of high-fidelity audio generative models and their potential for misuse. Given that audio large lang…
Manipulated Regions Localization For Partially Deepfake Audio: A Survey
Jiayi He, Jiangyan Yi, Jianhua Tao +2
With the development of audio deepfake techniques, attacks with partially deepfake audio are beginning to rise. Compared to fully deepfake, it is much harder to be identified by th…
StgcDiff: Spatial-Temporal Graph Condition Diffusion for Sign Language Transition Generation
Jiashu He, Jiayi He, Shengeng Tang +3
Sign language transition generation seeks to convert discrete sign language segments into continuous sign videos by synthesizing smooth transitions. However,most existing methods m…