Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Visual symbolic mechanisms: Emergent symbol processing in vision language models
Rim Assouel, Declan Campbell, Yoshua Bengio +1
To accurately process a visual scene, observers must bind features together to represent individual objects. This capacity is necessary, for instance, to distinguish an image conta…
cs.CV2025
Caption This, Reason That: VLMs Caught in the Middle
Zihan Weng, Lucas Gomez, Taylor Whittington Webb +1
Vision-Language Models (VLMs) have shown remarkable progress in visual understanding in recent years. Yet, they still lag behind human capabilities in specific visual tasks such as…