#prompt injection

topicprompt injection

9 papers · 1 filter

cs.CR2026

Piggybacking on Perception: Stealthy Concurrent Audio Prompt Injections against Multimodal LLM Agents

Mingxiao Liu, Yitong Li, Haoren Zhao +6

The paper studies stealthy audio prompt injection attacks that hide malicious instructions within normal speech to hijack multimodal LLM agents, introduces a benchmark (AudioAgentS…

cs.CR2026

GPT-Red: Automated Red Teaming via Self-Play at Scale

Eric Wallace, Christopher A. Choquette-Choo, Nikhil Kandpal +15

The paper presents GPT-Red, an automated red‑teaming system that uses self‑play to generate novel prompt‑injection attacks against large language models and improve their robustnes…

cs.CR2026

Bad Memory: Evaluating Prompt Injection Risks from Memory in Agentic Systems

Soham Gadgil, David Alexander, Sai Sunku +1

The paper investigates how malicious instructions embedded in persistent memory files can be used to launch prompt injection attacks on agentic AI systems, evaluating several large…

cs.CR2026

Context Contamination in LLM Analysis of Network Security Logs: Poison with Passive Prompt Injection and Mitigation Evaluation

Rabimba Karanjai, Yang Lu, Hemanth Hegadehalli Madhavarao +2

The paper shows that large language models used to analyze network security logs can be tricked by malicious log entries that inject hidden prompts, and it evaluates attacks and de…

cs.CR2026

Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavioral Objective Violation

Mohammad Allahbakhsh, Mohammad Hassan Bahari, Moslem Attar-Raouf

The paper proposes a new way to conduct penetration testing for AI-enabled systems by focusing on inducing undesirable AI-driven behavior that violates operational objectives, rath…

cs.CR2026

How Agents Ask for Permission: User Permissions for AI Agents, from Interfaces to Enforcement

Alexandra E. Michael, Franziska Roesner

The paper surveys how AI agents manage user‑level permission policies, builds a taxonomy of permission specification, derivation, and enforcement, and compares academic proposals w…