10 papers
UC-VLM: Consistency-Driven Learning for AI-Generated Image Detection with Vision-Language Large Models
Lei Tan, Shuwei Li, Mohan Kankanhalli +1
Vision-Language Large Models (VLLMs) are promising for AI-generated image (AIGI) detection because they can produce both a prediction and a natural-language output. However, most e…
FUSE: Frequency-domain Unification and Spectral Energy Alignment for Multi-modal Object Re-Identification
Xuanhao Qi, Tom H. Luan, Yukang Zhang +4
Despite significant progress in multi-modal Re-Identification (ReID), existing methods tend to emphasize low-frequency cues. Consequently, they focus on attributes such as color, i…
DPM++: Dynamic Masked Metric Learning for Occluded Person Re-identification
Lei Tan, Yingshi Luan, Pincong Zou +2
Although person re-identification has made impressive progress, occlusion caused by obstacles remains an unsettled issue in real applications. The difficulty lies in the mismatch b…
White-Balance First, Adjust Later: Cross-Camera Color Constancy via Vision-Language Evaluation
Shuwei Li, Lei Tan, Robby T. Tan
Color constancy aims to keep object colors consistent under varying illumination. Cross-camera generalization in color constancy remains challenging because learning-based models o…
Bridging Day and Night: Target-Class Hallucination Suppression in Unpaired Image Translation
Shuwei Li, Lei Tan, Robby T. Tan
Day-to-night unpaired image translation is important to downstream tasks but remains challenging due to large appearance shifts and the lack of direct pixel-level supervision. Exis…
Aggregating Diverse Cue Experts for AI-Generated Image Detection
Lei Tan, Shuwei Li, Mohan Kankanhalli +1
The rapid emergence of image synthesis models poses challenges to the generalization of AI-generated image detectors. However, existing methods often rely on model-specific feature…