2 citations · 3 across the 5 of their papers we have counts for
Showing 2024 · cs.CRShow all
2 papers · 2 filters
cs.CR2024
Virtual Context: Enhancing Jailbreak Attacks with Special Token Injection
Yuqi Zhou, Lin Lu, Hanchi Sun +2
Jailbreak attacks on large language models (LLMs) involve inducing these models to generate harmful content that violates ethics or laws, posing a significant threat to LLM securit…
cs.CR2024
AutoJailbreak: Exploring Jailbreak Attacks and Defenses through a Dependency Lens
Lin Lu, Hai Yan, Zenghui Yuan +4
Jailbreak attacks in large language models (LLMs) entail inducing the models to generate content that breaches ethical and legal norm through the use of malicious prompts, posing a…