Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
Xikang Yang, Biyu Zhou, Xuehai Tang +2
Large Language Models (LLMs) demonstrate impressive capabilities across a wide range of tasks, yet their safety mechanisms remain susceptible to adversarial attacks that exploit co…
cs.CL2024
Chain of Attack: a Semantic-Driven Contextual Multi-Turn attacker for LLM
Xikang Yang, Xuehai Tang, Songlin Hu +1
Large language models (LLMs) have achieved remarkable performance in various natural language processing tasks, especially in dialogue systems. However, LLM may also pose security…