45 citations · 84 across the 12 of their papers we have counts for
Showing 2022Show all
3 papers · 1 filter
cs.CV2022
Describing Sets of Images with Textual-PCA
Oded Hupert, Idan Schwartz, Lior Wolf
We seek to semantically describe a set of images, capturing both the attributes of single images and the variations within the set. Our procedure is analogous to Principle Componen…
cs.CV2022★ 14 cited
Zero-Shot Video Captioning with Evolving Pseudo-Tokens
Yoad Tewel, Yoav Shalev, Roy Nadler +2
We introduce a zero-shot video captioning method that employs two frozen networks: the GPT-2 language model and the CLIP image-text matching model. The matching score is used to st…
cs.CV2022★ 12 cited
Optimizing Relevance Maps of Vision Transformers Improves Robustness
Hila Chefer, Idan Schwartz, Lior Wolf
It has been observed that visual classification models often rely mostly on the image background, neglecting the foreground, which hurts their robustness to distribution changes. T…