1 paper · 1 filter
Yiming Zhang, Zicheng Zhang, Xinyi Wei +3
Current Visual Language Models (VLMs) show impressive image understanding but struggle with visual illusions, especially in real-world scenarios. Existing benchmarks focus on class…