2 papers
cs.AI2026
A Cross-Architecture Audit of Direction-Based Inference-Time Defences in Vision-Language Models
Xiangyu Yin, Tora Bodin, Rohan Menon +1
Inference time defences against vision language model jailbreaks often subtract a calibrated direction from the residual stream at a chosen decoder layer. We compare five defence c…
cs.LG2025
LipShiFT: A Certifiably Robust Shift-based Vision Transformer
Rohan Menon, Nicola Franco, Stephan Günnemann
Deriving tight Lipschitz bounds for transformer-based architectures presents a significant challenge. The large input sizes and high-dimensional attention modules typically prove t…