16 citations · 20 across the 29 of their papers we have counts for
9 papers · 1 filter
HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales
Wenbo Xu, Zhimin Chen, Xiaojie Liang +3
Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, presenting unprecedented challenges to…
NullEdit: Stealthy Image Protection via VLM Condition Redirection
Weiyao Huang, Liqin Wang, Ziqi Sheng +1
Modern image editors combine vision-language models (VLMs) with diffusion transformer backbones to modify a single reference image according to instructions without fine-tuning. Th…
I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models
Yimao Guo, Zuomin Qu, Wei Lu
The rapid advancement of video generation models has led to the increasing misuse of image-to-video (I2V) models. Although substantial progress has been made in detecting AI-genera…
CIEC: Coupling Implicit and Explicit Cues for Multimodal Weakly Supervised Manipulation Localization
Xinquan Yu, Wei Lu, Xiangyang Luo +1
To mitigate the threat of misinformation, multimodal manipulation localization has garnered growing attention. Consider that current methods rely on costly and time-consuming fine-…
Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection
Changjiang Jiang, Xinkuan Sha, Fengchang Yu +5
Recent studies have demonstrated that incorporating Chain-of-Thought (CoT) reasoning into the detection process can enhance a model's ability to detect synthetic images. However, e…
MARE: Multimodal Alignment and Reinforcement for Explainable Deepfake Detection via Vision-Language Models
Wenbo Xu, Wei Lu, Xiangyang Luo +1
Deepfake detection is a widely researched topic that is crucial for combating the spread of malicious content, with existing methods mainly modeling the problem as classification o…