3 papers
cs.CV2026
What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness
Yusheng He, Jizhe Zhou, Xia Du +3
Hallucination remains one of the key challenges undermining the reliability of Large Vision-Language Models (LVLMs). But what makes an LVLM hallucinate less? Many existing efforts…
cs.CV2025
Style Quantization for Data-Efficient GAN Training
Jian Wang, Xin Lan, Jizhe Zhou +2
Under limited data setting, GANs often struggle to navigate and effectively exploit the input latent space. Consequently, images generated from adjacent variables in a sparse input…
cs.CV2024
IMDL-BenCo: A Comprehensive Benchmark and Codebase for Image Manipulation Detection & Localization
Xiaochen Ma, Xuekang Zhu, Lei Su +8
A comprehensive benchmark is yet to be established in the Image Manipulation Detection & Localization (IMDL) field. The absence of such a benchmark leads to insufficient and mislea…