4 papers
Sequential Enumeration in Large Language Models
Kuinan Hou, Marco Zorzi, Alberto Testolin
Reliably counting and generating sequences of items remain a significant challenge for neural networks, including Large Language Models (LLMs). Indeed, although this capability is…
Assessing the Visual Enumeration Abilities of Specialized Counting Architectures and Vision-Language Models
Kuinan Hou, Jing Mi, Marco Zorzi +2
Counting the number of items in a visual scene remains a fundamental yet challenging task in computer vision. Traditional approaches to solving this problem rely on domain-specific…
Visual Enumeration Remains Challenging for Multimodal Generative AI
Alberto Testolin, Kuinan Hou, Marco Zorzi
Many animal species can approximately judge the number of objects in a visual scene at a single glance, and humans can further determine the exact cardinality of a set by deploying…
Estimating the distribution of numerosity and non-numerical visual magnitudes in natural scenes using computer vision
Kuinan Hou, Marco Zorzi, Alberto Testolin
Humans share with many animal species the ability to perceive and approximately represent the number of objects in visual scenes. This ability improves throughout childhood, sugges…