14 citations · 38 across the 11 of their papers we have counts for
1 paper · 2 filters
Paul Osemudiame Oamen, Owusu-Banahene Osei, Ananya Mukherjee +4
Existing vision-language model (VLM) benchmarks emphasize perception and reasoning accuracy (how well VLMs describe and reason about what they see in an image), with limited attent…