3 papers
cs.CV2026
Jailbreaking Vision-Language Models Through the Visual Modality
Aharon Azulay, Jan DubiÅski, Zhuoyun Li +2
The visual modality of vision-language models (VLMs) is an underexplored attack surface for bypassing safety alignment. We introduce four jailbreak attacks exploiting the vision co…
cs.CV2025
Revisiting CroPA: A Reproducibility Study and Enhancements for Cross-Prompt Adversarial Transferability in Vision-Language Models
Atharv Mittal, Agam Pandey, Amritanshu Tiwari +2
Large Vision-Language Models (VLMs) have revolutionized computer vision, enabling tasks such as image classification, captioning, and visual question answering. However, they remai…
cs.LG2024
LoRA Unlearns More and Retains More (Student Abstract)
Atharv Mittal
Due to increasing privacy regulations and regulatory compliance, Machine Unlearning (MU) has become essential. The goal of unlearning is to remove information related to a specific…