6 citations · 7 across the 2 of their papers we have counts for
1 paper · 1 filter
Weiwei Qi, Shuo Shao, Wei Gu +4
Large Language Models (LLMs) have exhibited remarkable capabilities but remain vulnerable to jailbreaking attacks, which can elicit harmful content from the models by manipulating…