1 paper
Guanyuan Pan, Shuai Wang, Yugui Lin +4
Vision Language Models (VLMs) have demonstrated remarkable potential in multimodal reasoning, yet they inherently suffer from spatial blindness and logical hallucinations when inte…