9 papers
SURE: Semi-dense Uncertainty-REfined Feature Matching
Sicheng Li, Zaiwang Gu, Jie Zhang +3
Establishing reliable image correspondences is essential for many robotic vision problems. However, existing methods often struggle in challenging scenarios with large viewpoint ch…
Visible Yet Unreadable: A Systematic Blind Spot of Vision Language Models Across Writing Systems
Jie Zhang, Ting Xu, Gelei Deng +5
Writing is a universal cultural technology that reuses vision for symbolic communication. Humans display striking resilience: we readily recognize words even when characters are fr…
IRCopilot: Automated Incident Response with Large Language Models
Xihuan Lin, Jie Zhang, Gelei Deng +4
Incident response plays a pivotal role in mitigating the impact of cyber attacks. In recent years, the intensity and complexity of global cyber threats have grown significantly, ma…
FaceTracer: Unveiling Source Identities from Swapped Face Images and Videos for Fraud Prevention
Zhongyi Zhang, Jie Zhang, Wenbo Zhou +5
Face-swapping techniques have advanced rapidly with the evolution of deep learning, leading to widespread use and growing concerns about potential misuse, especially in cases of fr…
Mask Image Watermarking
Runyi Hu, Jie Zhang, Shiqian Zhao +5
We present MaskWM, a simple, efficient, and flexible framework for image watermarking. MaskWM has two variants: (1) MaskWM-D, which supports global watermark embedding, watermark l…
SafeGuider: Robust and Practical Content Safety Control for Text-to-Image Models
Peigui Qi, Kunsheng Tang, Wenbo Zhou +5
Text-to-image models have shown remarkable capabilities in generating high-quality images from natural language descriptions. However, these models are highly vulnerable to adversa…