From the 2 of 73 linked papers with an AI index.
5 papers · 1 filter
Diffusion-based Cumulative Adversarial Purification for Vision Language Models
Jia Fu, Yongtao Wu, Yihang Chen +5
Vision Language Models (VLMs) have shown remarkable capabilities in multimodal understanding, yet their susceptibility to adversarial perturbations poses a significant threat to th…
Spatial Priors via Space Filling Curves for Small and Limited Data Vision Transformers
Leyla Naz Candogan, Arshia Afzal, Pol Puigdemont +1
Though Vision Transformers (ViTs) have become the dominant backbone in many computer vision tasks, due to permutation equivariance, their attention mechanism lacks explicit spatial…
The Last Mile to Supervised Performance: Semi-Supervised Domain Adaptation for Semantic Segmentation
Daniel Morales-Brotons, Grigorios Chrysos, Stratis Tzoumas +1
Supervised deep learning requires massive labeled datasets, but obtaining annotations is not always easy or possible, especially for dense tasks like semantic segmentation. To over…
Membership Inference Attacks against Large Vision-Language Models
Zhan Li, Yongtao Wu, Yihang Chen +3
Large vision-language models (VLLMs) exhibit promising capabilities for processing multi-modal tasks across various application scenarios. However, their emergence also raises sign…
Going beyond Compositions, DDPMs Can Produce Zero-Shot Interpolations
Justin Deschenaux, Igor Krawczuk, Grigorios Chrysos +1
Denoising Diffusion Probabilistic Models (DDPMs) exhibit remarkable capabilities in image generation, with studies suggesting that they can generalize by composing latent factors l…