3 papers
cs.CV2026
Stage-wise Attention-Guided Region Sequencing for Adversarial Attacks on Large Vision-Language Models
Jaehyun Kwak, Nam Cao, Boryeong Cho +3
Targeted adversarial attacks on Large Vision-Language Models (LVLMs) test whether small image perturbations can steer model responses toward attacker-specified content. Under the s…
cs.LG2026
Generative Visual Code Mobile World Models
Woosung Koh, Sungjun Han, Segyu Lee +2
Mobile Graphical User Interface (GUI) World Models (WMs) offer a promising path for improving mobile GUI agent performance at train- and inference-time. However, current approaches…
cs.CV2026
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
Segyu Lee, Boryeong Cho, Hojung Jung +8
Unified Multimodal Models (UMMs) offer powerful cross-modality capabilities but introduce new safety risks not observed in single-task models. Despite their emergence, existing saf…