backdoor attacks 1black-box attacks 1data poisoning 1representation alignment 1trigger generalization 1
From the 1 of 29 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
ACIArena: Toward Unified Evaluation for Agent Cascading Injection
Hengyu An, Minxi Li, Jinghuai Zhang +6
Collaboration and information sharing empower Multi-Agent Systems (MAS) but also introduce a critical security risk known as Agent Cascading Injection (ACI). In such attacks, a com…
cs.AI2024
Navigating the Risks: A Survey of Security, Privacy, and Ethics Threats in LLM-Based Agents
Yuyou Gan, Yong Yang, Zhe Ma +10
With the continuous development of large language models (LLMs), transformer-based models have made groundbreaking advances in numerous natural language processing (NLP) tasks, lea…