8 papers · 1 filter
NAMESAKES: Probing Identity Memorization in Text-to-Image Models
Morris Alper, Vasudha Varadarajan, Moran Yanuka +2
Text-to-image (T2I) models generate realistic likenesses of some individuals when prompted with their names, raising privacy concerns. However, distinguishing whether a generated f…
WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild
Morris Alper, David Novotny, Filippos Kokkinos +2
Despite recent advances in sparse novel view synthesis (NVS) applied to object-centric scenes, scene-level NVS remains a challenge. A central issue is the lack of available clean m…
Dynamic Scene Understanding from Vision-Language Representations
Shahaf Pruss, Morris Alper, Hadar Averbuch-Elor
Images depicting complex, dynamic scenes are challenging to parse automatically, requiring both high-level comprehension of the overall situation and fine-grained identification of…
ProtoSnap: Prototype Alignment for Cuneiform Signs
Rachel Mikulinsky, Morris Alper, Shai Gordin +3
The cuneiform writing system served as the medium for transmitting knowledge in the ancient Near East for a period of over three thousand years. Cuneiform signs have a complex inte…
WAFFLE: Multimodal Floorplan Understanding in the Wild
Keren Ganon, Morris Alper, Rachel Mikulinsky +1
Buildings are a central feature of human culture and are increasingly being analyzed with computational methods. However, recent works on computational building understanding have…
Mitigating Open-Vocabulary Caption Hallucinations
Assaf Ben-Kish, Moran Yanuka, Morris Alper +2
While recent years have seen rapid progress in image-conditioned text generation, image captioning still suffers from the fundamental issue of hallucinations, namely, the generatio…