Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
AgentCanary: A Security Evaluation Framework for Autonomous AI Agents in Real Executable Environments
Peiyang Li, Songping Wang, Yi Huang +9
Autonomous AI agents have driven the transition from conversation to task execution, shifting security failures from textual deception to system compromise. Although security evalu…
cs.CR2026
Confundo: Learning to Generate Robust Poison for Practical RAG Systems
Haoyang Hu, Zhejun Jiang, Yueming Lyu +3
Retrieval-augmented generation (RAG) is increasingly deployed in real-world applications, where its reference-grounded design makes outputs appear trustworthy. This trust has spurr…