5 citations · 5 across the 4 of their papers we have counts for
4 papers
SAGE: Saliency-Guided Mixup with Optimal Rearrangements
Avery Ma, Nikita Dvornik, Ran Zhang +3
Data augmentation is a key element for training accurate models by reducing overfitting and improving generalization. For image classification, the most popular data augmentation t…
Visual Semantic Parsing: From Images to Abstract Meaning Representation
Mohamed Ashraf Abdelsalam, Zhan Shi, Federico Fancellu +4
The success of scene graphs for visual scene understanding has brought attention to the benefits of abstracting a visual input (e.g., image) into a structured representation, where…
Uncertainty-based Cross-Modal Retrieval with Probabilistic Representations
Leila Pishdad, Ran Zhang, Konstantinos G. Derpanis +2
Probabilistic embeddings have proven useful for capturing polysemous word meanings, as well as ambiguity in image matching. In this paper, we study the advantages of probabilistic…
VASTA: A Vision and Language-assisted Smartphone Task Automation System
Alborz Rezazadeh Sereshkeh, Gary Leung, Krish Perumal +4
We present VASTA, a novel vision and language-assisted Programming By Demonstration (PBD) system for smartphone task automation. Development of a robust PBD automation system requi…