1 paper · 1 filter
Kejia Chen, Jiawen Zhang, Jiacong Hu +4
Vision-Language Models (VLMs) have become essential backbones of modern multimodal intelligence, yet their outputs remain prone to hallucination-plausible text misaligned with visu…