2 papers
cs.CV2026
Robust MLLM Unlearning via Visual Knowledge Distillation
Yuhang Wang, Zhenxing Niu, Haoxuan Ji +3
Recently, machine unlearning approaches have been proposed to remove sensitive information from well-trained large models. However, most existing methods are tailored for LLMs, whi…
cs.AI2025
Efficient LLM-Jailbreaking via Multimodal-LLM Jailbreak
Haoxuan Ji, Zheng Lin, Zhenxing Niu +2
This paper focuses on jailbreaking attacks against large language models (LLMs), eliciting them to generate objectionable content in response to harmful user queries. Unlike previo…