1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Jiwei Guan, Tianyu Ding, Longbing Cao +3
Vision-language pretraining (VLP) with transformers has demonstrated exceptional performance across numerous multimodal tasks. However, the adversarial robustness of these models h…