1 paper
Taeyun Roh, Suhyeong Park, Eun-yeong Jo +5
Multimodal multiple-choice question answering (MCQA) provides a standardized and objectively measurable setting for evaluating vision-language models (VLMs). However, because the M…