#prompt injection
9 papers · 1 filter
Piggybacking on Perception: Stealthy Concurrent Audio Prompt Injections against Multimodal LLM Agents
Mingxiao Liu, Yitong Li, Haoren Zhao +6
The paper studies stealthy audio prompt injection attacks that hide malicious instructions within normal speech to hijack multimodal LLM agents, introduces a benchmark (AudioAgentS…
GPT-Red: Automated Red Teaming via Self-Play at Scale
Eric Wallace, Christopher A. Choquette-Choo, Nikhil Kandpal +15
The paper presents GPT-Red, an automated red‑teaming system that uses self‑play to generate novel prompt‑injection attacks against large language models and improve their robustnes…
Bad Memory: Evaluating Prompt Injection Risks from Memory in Agentic Systems
Soham Gadgil, David Alexander, Sai Sunku +1
The paper investigates how malicious instructions embedded in persistent memory files can be used to launch prompt injection attacks on agentic AI systems, evaluating several large…
Context Contamination in LLM Analysis of Network Security Logs: Poison with Passive Prompt Injection and Mitigation Evaluation
Rabimba Karanjai, Yang Lu, Hemanth Hegadehalli Madhavarao +2
The paper shows that large language models used to analyze network security logs can be tricked by malicious log entries that inject hidden prompts, and it evaluates attacks and de…
Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavioral Objective Violation
Mohammad Allahbakhsh, Mohammad Hassan Bahari, Moslem Attar-Raouf
The paper proposes a new way to conduct penetration testing for AI-enabled systems by focusing on inducing undesirable AI-driven behavior that violates operational objectives, rath…
How Agents Ask for Permission: User Permissions for AI Agents, from Interfaces to Enforcement
Alexandra E. Michael, Franziska Roesner
The paper surveys how AI agents manage user‑level permission policies, builds a taxonomy of permission specification, derivation, and enforcement, and compares academic proposals w…