7 citations · 7 across the 3 of their papers we have counts for
1 paper · 1 filter
Kesheng Chen, Yamin Hu, Qi Zhou +2
Vision-language models (VLMs) achieve strong performance on many benchmarks, yet a basic reliability question remains underexplored: when visual evidence conflicts with commonsense…