2 papers
cs.CV2026
Jailbreaking Vision-Language Models Through the Visual Modality
Aharon Azulay, Jan DubiÅski, Zhuoyun Li +2
The visual modality of vision-language models (VLMs) is an underexplored attack surface for bypassing safety alignment. We introduce four jailbreak attacks exploiting the vision co…
cs.CV2025
V-LASIK: Consistent Glasses-Removal from Videos Using Synthetic Data
Rotem Shalev-Arkushin, Aharon Azulay, Tavi Halperin +3
Diffusion-based generative models have recently shown remarkable image and video editing capabilities. However, local video editing, particularly removal of small attributes like g…