13 papers
Cultivating Forensic Reasoning for Generalizable Multimodal Manipulation Detection
Yuchen Zhang, Yaxiong Wang, Kecheng Han +4
Recent advances in generative AI have significantly enhanced the realism of multimodal media manipulation, thereby posing substantial challenges to manipulation detection. Existing…
Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search
Jiahao Zhang, Shaofei Huang, Yaxiong Wang +1
Text-based person search faces inherent limitations due to data scarcity, driven by stringent privacy constraints and the high cost of manual annotation. To mitigate this, existing…
Generating Attribution Reports for Manipulated Facial Images: A Dataset and Baseline
Jingchun Lian, Lingyu Liu, Yaxiong Wang +4
Existing facial forgery detection methods typically focus on binary classification or pixel-level localization, providing little semantic insight into the nature of the manipulatio…
SEED: A Large-Scale Benchmark for Provenance Tracing in Sequential Deepfake Facial Edits
Mengieong Hoi, Zhedong Zheng, Ping Liu +1
Deepfake content on social networks is increasingly produced through multiple \emph{sequential} edits to biometric data such as facial imagery. Consequently, the final appearance o…
Harnessing Weak Pair Uncertainty for Text-based Person Search
Jintao Sun, Zhedong Zheng, Gangyi Ding
In this paper, we study the text-based person search, which is to retrieve the person of interest via natural language description. Prevailing methods usually focus on the strict o…
Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation
Lingyu Liu, Yaxiong Wang, Li Zhu +1
Video frame interpolation aims to synthesize realistic intermediate frames between given endpoints while adhering to specific motion semantics. While recent generative models have…