16 citations · 18 across the 11 of their papers we have counts for
1 paper · 1 filter
Peiyan Zhang, Haibo Jin, Liying Kang +1
Jailbreak attacks reveal critical vulnerabilities in Large Language Models (LLMs) by causing them to generate harmful or unethical content. Evaluating these threats is particularly…