40 citations · 73 across the 14 of their papers we have counts for
3 papers · 1 filter
The Geometry of Representational Failures in Vision Language Models
Daniele Savietto, Declan Campbell, André Panisson +4
Vision-Language Models (VLMs) exhibit puzzling failures in multi-object visual tasks, such as hallucinating non-existent elements or failing to identify the most similar objects am…
HOLMES: HOLonym-MEronym based Semantic inspection for Convolutional Image Classifiers
Francesco Dibitonto, Fabio Garcea, André Panisson +2
Convolutional Neural Networks (CNNs) are nowadays the model of choice in Computer Vision, thanks to their ability to automatize the feature extraction process in visual tasks. Howe…
Beyond One-Hot-Encoding: Injecting Semantics to Drive Image Classifiers
Alan Perotti, Simone Bertolotto, Eliana Pastor +1
Images are loaded with semantic information that pertains to real-world ontologies: dog breeds share mammalian similarities, food pictures are often depicted in domestic environmen…