1 citations · 2 across the 18 of their papers we have counts for
1 paper · 1 filter
Tianyi Wu, Zhiwei Xue, Yue Liu +3
Jailbreak attacks, which aim to cause LLMs to perform unrestricted behaviors, have become a critical and challenging direction in AI safety. Despite achieving the promising attack…