From the 1 of 27 linked papers with an AI index.
27 papers
NullEdit: Stealthy Image Protection via VLM Condition Redirection
Weiyao Huang, Liqin Wang, Ziqi Sheng +1
Modern image editors combine vision-language models (VLMs) with diffusion transformer backbones to modify a single reference image according to instructions without fine-tuning. Th…
I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models
Yimao Guo, Zuomin Qu, Wei Lu
The paper introduces I2VShield, a lightweight proactive defense that generates text‑adaptive perturbations and uses a multimodal attention disruption attack to protect images from…
HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales
Wenbo Xu, Zhimin Chen, Xiaojie Liang +3
Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, presenting unprecedented challenges to…
Mining Forgery Traces from Reconstruction Error: A Weakly Supervised Framework for Multimodal Deepfake Temporal Localization
Midou Guo, Qilin Yin, Wei Lu +1
Modern deepfakes have evolved into localized and intermittent manipulations that require fine-grained temporal localization to mitigate severe digital security risks. The prohibiti…
Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection
Changjiang Jiang, Xinkuan Sha, Fengchang Yu +5
Recent studies have demonstrated that incorporating Chain-of-Thought (CoT) reasoning into the detection process can enhance a model's ability to detect synthetic images. However, e…
FakeScope: Large Multimodal Expert Model for Transparent AI-Generated Image Forensics
Yixuan Li, Yu Tian, Yipo Huang +4
The rapid and unrestrained advancement of generative artificial intelligence (AI) presents a double-edged sword. While enabling unprecedented creativity, it also facilitates the ge…