20 citations · 33 across the 7 of their papers we have counts for
3 papers · 1 filter
Caption supervision enables robust learners
Benjamin Feuer, Ameya Joshi, Chinmay Hegde
Vision language (VL) models like CLIP are robust to natural distribution shifts, in part because CLIP learns on unstructured data using a technique called caption supervision; the…
Adversarial Token Attacks on Vision Transformers
Ameya Joshi, Gauri Jagatap, Chinmay Hegde
Vision transformers rely on a patch token based self attention mechanism, in contrast to convolutional networks. We investigate fundamental differences between these two families o…
Semantic Adversarial Attacks: Parametric Transformations That Fool Deep Classifiers
Ameya Joshi, Amitangshu Mukherjee, Soumik Sarkar +1
Deep neural networks have been shown to exhibit an intriguing vulnerability to adversarial input images corrupted with imperceptible perturbations. However, the majority of adversa…