5 papers · 1 filter
ALLUDE: A Unified Evaluation System for Configurable Attacks in Differentiable Environments
Mansi Phute, Alexander Greenhalgh, Matthew Hull +8
Adversarial attacks against vision models like object detectors are often evaluated under limited conditions, leaving their performance under-characterized. Bridging simulation and…
VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models
Ravikumar Balakrishnan, Mansi Phute
As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has become increasingly important. Existi…
VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models
Mansi Phute, Ravikumar Balakrishnan
Vision Language Models (VLMs) are increasingly being used in a broad range of applications, bringing their security and behavioral control to the forefront. While existing approach…
ComplicitSplat: Downstream Models are Vulnerable to Blackbox Attacks by 3D Gaussian Splat Camouflages
Matthew Hull, Haoyang Yang, Pratham Mehta +8
As 3D Gaussian Splatting (3DGS) gains rapid adoption in safety-critical tasks for efficient novel-view synthesis from static images, how might an adversary tamper images to cause h…
Semi-Truths: A Large-Scale Dataset of AI-Augmented Images for Evaluating Robustness of AI-Generated Image detectors
Anisha Pal, Julia Kruk, Mansi Phute +4
Text-to-image diffusion models have impactful applications in art, design, and entertainment, yet these technologies also pose significant risks by enabling the creation and dissem…