2 papers
cs.CR2026
Unreal Thinking: Chain-of-Thought Hijacking via Two-stage Backdoor
Wenhan Chang, Tianqing Zhu, Ping Xiong +2
Large Language Models (LLMs) are increasingly deployed in settings where Chain-of-Thought (CoT) is interpreted by users. This creates a new safety risk: attackers may manipulate th…
cs.CR2026
Chain-of-Lure: A Universal Jailbreak Attack Framework using Unconstrained Synthetic Narratives
Wenhan Chang, Tianqing Zhu, Yu Zhao +3
In the era of rapid generative AI development, interactions with large language models (LLMs) pose increasing risks of misuse. Prior research has primarily focused on attacks using…