5 papers
Detecting Lip-Syncing Deepfakes: Vision Temporal Transformer for Analyzing Mouth Inconsistencies
Soumyya Kanti Datta, Shan Jia, Siwei Lyu
Deepfakes are AI-generated media in which the original content is digitally altered to create convincing but manipulated images, videos, or audio. Among the various types of deepfa…
Dense Feature Interaction Network for Image Inpainting Localization
Ye Yao, Tingfeng Han, Shan Jia +1
Image inpainting, the process of filling in missing areas in an image, is a common image editing technique. Inpainting can be used to conceal or alter image contents in malicious m…
ParallelEdits: Efficient Multi-object Image Editing
Mingzhen Huang, Jialing Cai, Shan Jia +2
Text-driven image synthesis has made significant advancements with the development of diffusion models, transforming how visual content is generated from text prompts. Despite thes…
Explicit Correlation Learning for Generalizable Cross-Modal Deepfake Detection
Cai Yu, Shan Jia, Xiaomeng Fu +6
With the rising prevalence of deepfakes, there is a growing interest in developing generalizable detection methods for various types of deepfakes. While effective in their specific…
Exposing Text-Image Inconsistency Using Diffusion Models
Mingzhen Huang, Shan Jia, Zhou Zhou +3
In the battle against widespread online misinformation, a growing problem is text-image inconsistency, where images are misleadingly paired with texts with different intent or mean…