most citedAI-Synthesized Voice Detection Using Neural Vocoder Artifacts

4 citations · 9 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CV2024

Explicit Correlation Learning for Generalizable Cross-Modal Deepfake Detection

Cai Yu, Shan Jia, Xiaomeng Fu +6

With the rising prevalence of deepfakes, there is a growing interest in developing generalizable detection methods for various types of deepfakes. While effective in their specific…

cs.CV2024

Exposing Text-Image Inconsistency Using Diffusion Models

Mingzhen Huang, Shan Jia, Zhou Zhou +3

In the battle against widespread online misinformation, a growing problem is text-image inconsistency, where images are misleadingly paired with texts with different intent or mean…

cs.CV20233 cited

Integrating Audio-Visual Features for Multimodal Deepfake Detection

Sneha Muppalla, Shan Jia, Siwei Lyu

Deepfakes are AI-generated media in which an image or video has been digitally modified. The advancements made in deepfake technology have led to privacy and security issues. Most…

cs.CV2023

UPDExplainer: an Interpretable Transformer-based Framework for Urban Physical Disorder Detection Using Street View Imagery

Chuanbo Hu, Shan Jia, Fan Zhang +4

Urban Physical Disorder (UPD), such as old or abandoned buildings, broken sidewalks, litter, and graffiti, has a negative impact on residents' quality of life. They can also increa…

cs.SD20234 cited

AI-Synthesized Voice Detection Using Neural Vocoder Artifacts

Chengzhe Sun, Shan Jia, Shuwei Hou +1

Advancements in AI-synthesized human voices have created a growing threat of impersonation and disinformation, making it crucial to develop methods to detect synthetic human voices…

cs.CV20232 cited

AutoSplice: A Text-prompt Manipulated Image Dataset for Media Forensics

Shan Jia, Mingzhen Huang, Zhou Zhou +3

Recent advancements in language-image models have led to the development of highly realistic images that can be generated from textual descriptions. However, the increased visual q…