1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Sunny Gupta, Shounak Das, Amit Sethi
Vision language foundation models such as CLIP exhibit impressive zero-shot generalization yet remain vulnerable to spurious correlations across visual and textual modalities. Exis…