1 paper · 1 filter
Amelia F. Hardy, Houjun Liu, Allie Griffith +3
Existing LLM red-teaming approaches prioritize high attack success rate, often resulting in high-perplexity prompts. This focus overlooks low-perplexity attacks that are more diffi…