2 papers
cs.CR2024
Global Challenge for Safe and Secure LLMs Track 1
Xiaojun Jia, Yihao Huang, Yang Liu +27
This paper introduces the Global Challenge for Safe and Secure Large Language Models (LLMs), a pioneering initiative organized by AI Singapore (AISG) and the CyberSG R&D Programme…
cs.AI2024
Boosting Jailbreak Transferability for Large Language Models
Hanqing Liu, Lifeng Zhou, Huanqian Yan
Large language models have drawn significant attention to the challenge of safe alignment, especially regarding jailbreak attacks that circumvent security measures to produce harmf…