Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Adaptive Forensic Feature Refinement via Intrinsic Importance Perception
Jiazhen Yang, Junjun Zheng, Kejia Chen +5
With the rapid development of generative models and multimodal content editing technologies, the key challenge faced by synthetic image detection (SID) lies in cross-distribution g…
cs.CV2025
Token-Level Inference-Time Alignment for Vision-Language Models
Kejia Chen, Jiawen Zhang, Jiacong Hu +4
Vision-Language Models (VLMs) have become essential backbones of modern multimodal intelligence, yet their outputs remain prone to hallucination-plausible text misaligned with visu…
cs.CV2025
SHAPE : Self-Improved Visual Preference Alignment by Iteratively Generating Holistic Winner
Kejia Chen, Jiawen Zhang, Jiacong Hu +4
Large Visual Language Models (LVLMs) increasingly rely on preference alignment to ensure reliability, which steers the model behavior via preference fine-tuning on preference data…