7 citations · 7 across the 1 of their papers we have counts for
4 papers · 1 filter
Dense and Aligned Captions (DAC) Promote Compositional Reasoning in VL Models
Sivan Doveh, Assaf Arbelle, Sivan Harary +9
Vision and Language (VL) models offer an effective method for aligning representation spaces of images and text, leading to numerous applications such as cross-modal retrieval, vis…
Detector-Free Weakly Supervised Grounding by Separation
Assaf Arbelle, Sivan Doveh, Amit Alfassy +14
Nowadays, there is an abundance of data involving images and surrounding free-form text weakly corresponding to those images. Weakly Supervised phrase-Grounding (WSG) deals with th…
StarNet: towards Weakly Supervised Few-Shot Object Detection
Leonid Karlinsky, Joseph Shtok, Amit Alfassy +8
Few-shot detection and classification have advanced significantly in recent years. Yet, detection approaches require strong annotation (bounding boxes) both for pre-training and fo…
LaSO: Label-Set Operations networks for multi-label few-shot learning
Amit Alfassy, Leonid Karlinsky, Amit Aides +5
Example synthesis is one of the leading methods to tackle the problem of few-shot learning, where only a small number of samples per class are available. However, current synthesis…