2 papers
cs.AI2026
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message
Wei Duan, Li Qian
The rise of conversational interfaces has greatly enhanced LLM usability by leveraging dialogue history for sophisticated reasoning. However, this reliance introduces an unexplored…
cs.LG2025
Dual-View Inference Attack: Machine Unlearning Amplifies Privacy Exposure
Lulu Xue, Shengshan Hu, Linqiang Qian +6
Machine unlearning is a newly popularized technique for removing specific training data from a trained model, enabling it to comply with data deletion requests. While it protects t…