12 citations · 27 across the 5 of their papers we have counts for
4 papers · 1 filter
A model for full local image interpretation
Guy Ben-Yosef, Liav Assif, Daniel Harari +1
We describe a computational model of humans' ability to provide a detailed interpretation of components in a scene. Humans can identify in an image meaningful components almost eve…
Image interpretation by iterative bottom-up top-down processing
Shimon Ullman, Liav Assif, Alona Strugatski +4
Scene understanding requires the extraction and representation of scene components together with their properties and inter-relations. We describe a model in which meaningful scene…
Detector-Free Weakly Supervised Grounding by Separation
Assaf Arbelle, Sivan Doveh, Amit Alfassy +14
Nowadays, there is an abundance of data involving images and surrounding free-form text weakly corresponding to those images. Weakly Supervised phrase-Grounding (WSG) deals with th…
What can human minimal videos tell us about dynamic recognition models?
Guy Ben-Yosef, Gabriel Kreiman, Shimon Ullman
In human vision objects and their parts can be visually recognized from purely spatial or purely temporal information but the mechanisms integrating space and time are poorly under…