3 papers
cs.AI2025
Visual serial processing deficits explain divergences in human and VLM reasoning
Nicholas Budny, Kia Ghods, Declan Campbell +6
Why do Vision Language Models (VLMs), despite success on standard benchmarks, often fail to match human performance on surprisingly simple visual reasoning tasks? While the underly…
cs.LG2025
Bound by semanticity: universal laws governing the generalization-identification tradeoff
Marco Nurisso, Jesseba Fernando, Raj Deshpande +9
Intelligent systems must deploy internal representations that are simultaneously structured -- to support broad generalization -- and selective -- to preserve input identity. We ex…
cs.AI2024
Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem
Declan Campbell, Sunayana Rane, Tyler Giallanza +8
Recent work has documented striking heterogeneity in the performance of state-of-the-art vision language models (VLMs), including both multimodal language models and text-to-image…