works on

From the 2 of 73 linked papers with an AI index.

activity
20242026
collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

Diffusion-based Cumulative Adversarial Purification for Vision Language Models

Jia Fu, Yongtao Wu, Yihang Chen +5

Vision Language Models (VLMs) have shown remarkable capabilities in multimodal understanding, yet their susceptibility to adversarial perturbations poses a significant threat to th…

cs.CV2026

Spatial Priors via Space Filling Curves for Small and Limited Data Vision Transformers

Leyla Naz Candogan, Arshia Afzal, Pol Puigdemont +1

Though Vision Transformers (ViTs) have become the dominant backbone in many computer vision tasks, due to permutation equivariance, their attention mechanism lacks explicit spatial…

cs.CV2024

The Last Mile to Supervised Performance: Semi-Supervised Domain Adaptation for Semantic Segmentation

Daniel Morales-Brotons, Grigorios Chrysos, Stratis Tzoumas +1

Supervised deep learning requires massive labeled datasets, but obtaining annotations is not always easy or possible, especially for dense tasks like semantic segmentation. To over…

cs.CV2024

Membership Inference Attacks against Large Vision-Language Models

Zhan Li, Yongtao Wu, Yihang Chen +3

Large vision-language models (VLLMs) exhibit promising capabilities for processing multi-modal tasks across various application scenarios. However, their emergence also raises sign…

cs.CV2024

Going beyond Compositions, DDPMs Can Produce Zero-Shot Interpolations

Justin Deschenaux, Igor Krawczuk, Grigorios Chrysos +1

Denoising Diffusion Probabilistic Models (DDPMs) exhibit remarkable capabilities in image generation, with studies suggesting that they can generalize by composing latent factors l…