2 papers
cs.CR2024
Improved Generation of Adversarial Examples Against Safety-aligned LLMs
Qizhang Li, Yiwen Guo, Wangmeng Zuo +1
Adversarial prompts generated using gradient-based methods exhibit outstanding performance in performing automatic jailbreak attacks against safety-aligned LLMs. Nevertheless, due…
cs.CR2024
Intrusion Detection at Scale with the Assistance of a Command-line Language Model
Jiongliang Lin, Yiwen Guo, Hao Chen
Intrusion detection is a long standing and crucial problem in security. A system capable of detecting intrusions automatically is on great demand in enterprise security solutions.…