fine-tuning robustness 1large language models 1margin calibration 1model unlearning 1relearn attacks 1
From the 1 of 5 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Fragile by Design: On the Limits of Adversarial Defenses in Personalized Generation
Zhen Chen, Yi Zhang, Xiangyu Yin +4
Personalized AI applications such as DreamBooth enable the generation of customized content from user images, but also raise significant privacy concerns, particularly the risk of…
cs.CV2025
TAIJI: Textual Anchoring for Immunizing Jailbreak Images in Vision Language Models
Xiangyu Yin, Yi Qi, Jinwei Hu +5
Vision Language Models (VLMs) have demonstrated impressive inference capabilities, but remain vulnerable to jailbreak attacks that can induce harmful or unethical responses. Existi…
cs.CV2025
CeTAD: Towards Certified Toxicity-Aware Distance in Vision Language Models
Xiangyu Yin, Jiaxu Liu, Zhen Chen +4
Recent advances in large vision-language models (VLMs) have demonstrated remarkable success across a wide range of visual understanding tasks. However, the robustness of these mode…