468 citations · 492 across the 28 of their papers we have counts for
1 paper · 1 filter
Martin Kuo, Jianyi Zhang, Aolin Ding +12
Malicious attackers can exploit large language models (LLMs) by engaging them in multi-turn dialogues to achieve harmful objectives, posing significant safety risks to society. To…