1 paper
Tao Huang, Rui Wang, Xiaofei Liu +3
%Large vision-language models (LVLMs) have shown substantial advances in multimodal understanding and generation. However, when presented with incompetent or adversarial inputs, th…