1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Daniele Savietto, Declan Campbell, André Panisson +4
Vision-Language Models (VLMs) exhibit puzzling failures in multi-object visual tasks, such as hallucinating non-existent elements or failing to identify the most similar objects am…