1 paper
Yu Yan, Sheng Sun, Shengjia Cheng +3
Vision-Language Models (VLMs) with multimodal reasoning capabilities are high-value attack targets, given their potential for handling complex multimodal harmful tasks. Mainstream…