9 papers
FusionAgent: A Multimodal Agent with Dynamic Model Selection for Human Recognition
Jie Zhu, Xiao Guo, Yiyang Su +2
Model fusion is a key strategy for robust recognition in unconstrained scenarios, as different models provide complementary strengths. This is especially important for whole-body h…
NAU-QMUL: Utilizing BERT and CLIP for Multi-modal AI-Generated Image Detection
Xiaoyu Guo, Arkaitz Zubiaga
With the aim of detecting AI-generated images and identifying the specific models responsible for their generation, we propose a multi-modal multi-task model. The model leverages p…
TalkingHeadBench: A Multi-Modal Benchmark & Analysis of Talking-Head DeepFake Detection
Xinqi Xiong, Prakrut Patel, Qingyuan Fan +6
The rapid advancement of talking-head deepfake generation fueled by advanced generative models has elevated the realism of synthetic videos to a level that poses substantial risks…
On the Holistic Approach for Detecting Human Image Forgery
Xiao Guo, Jie Zhu, Anil Jain +1
The rapid advancement of AI-generated content (AIGC) has escalated the threat of deepfakes, from facial manipulations to the synthesis of entire photorealistic human bodies. Howeve…
Benchmarking Unified Face Attack Detection via Hierarchical Prompt Tuning
Ajian Liu, Haocheng Yuan, Xiao Guo +13
PAD and FFD are proposed to protect face data from physical media-based Presentation Attacks and digital editing-based DeepFakes, respectively. However, isolated training of these…
On Learning Multi-Modal Forgery Representation for Diffusion Generated Video Detection
Xiufeng Song, Xiao Guo, Jiache Zhang +5
Large numbers of synthesized videos from diffusion models pose threats to information security and authenticity, leading to an increasing demand for generated content detection. Ho…