6 citations · 9 across the 37 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
Rosiana Natalie, Wenqian Xu, Ruei-Che Chang +2
Advances in vision language models (VLMs) have enabled the simulation of general human behavior through their reasoning and problem solving capabilities. However, prior research ha…
cs.CV2024
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
Longju Bai, Angana Borah, Oana Ignat +1
Large Multimodal Models (LMMs) exhibit impressive performance across various multimodal tasks. However, their effectiveness in cross-cultural contexts remains limited due to the pr…