1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Kai Hu, Weichen Yu, Yining Li +7
Recent research indicates that large language models (LLMs) are susceptible to jailbreaking attacks that can generate harmful content. This paper introduces a novel token-level att…