adversarial perturbations 1biometric security 1computer vision 1handwriting authentication 1style protection 1
From the 1 of 33 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection
Shuyu Jiang, Kaiyu Xu, Xingshu Chen +5
Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in dominant languages and has not pro…
cs.CL2025
Watch Out for Your Guidance on Generation! Exploring Conditional Backdoor Attacks against Large Language Models
Jiaming He, Wenbo Jiang, Guanyu Hou +3
Mainstream backdoor attacks on large language models (LLMs) typically set a fixed trigger in the input instance and specific responses for triggered queries. However, the fixed tri…