2 papers
cs.AI2026
On Protecting Agentic Systems' Intellectual Property via Watermarking
Liwen Wang, Zongjie Li, Yuchong Xie +4
The evolution of Large Language Models (LLMs) into agentic systems that perform autonomous reasoning and tool use has created significant intellectual property (IP) value. We demon…
cs.CR2025
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner
Xunguang Wang, Daoyuan Wu, Zhenlan Ji +7
Jailbreaking is an emerging adversarial attack that bypasses the safety alignment deployed in off-the-shelf large language models (LLMs) and has evolved into multiple categories: h…