1 paper · 1 filter
Jihyoung Jang, Hyounghun Kim
Visual Question Answering (VQA) is a core task for evaluating the capabilities of Vision-Language Models (VLMs). Existing VQA benchmarks primarily feature clear and unambiguous ima…