1 paper
Paul Osemudiame Oamen, Owusu-Banahene Osei, Ananya Mukherjee +4
Existing vision-language model (VLM) benchmarks emphasize perception and reasoning accuracy (how well VLMs describe and reason about what they see in an image), with limited attent…