2 papers
cs.CR2026
A Survey on Long-Term Memory Security in LLM Agents: Attacks, Defenses, and Governance Across the Memory Lifecycle
Zehao Lin, Xixuan Hao, Renyu Fu +5
The emergence of writable, cross-session persistent memory in LLM agents introduces a qualitatively different threat landscape from conventional input-centric security concerns, ch…
cs.CL2024
Playing Language Game with LLMs Leads to Jailbreaking
Yu Peng, Zewen Long, Fangming Dong +3
The advent of large language models (LLMs) has spurred the development of numerous jailbreak techniques aimed at circumventing their security defenses against malicious attacks. An…