3 papers
cs.CR2026
The Anatomy of a Prompt Injection: A Component Model for Structured Analysis
Jeremy McHugh
Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured exploits, despite advancing ag…
cs.CR2025
Prompt Injection 2.0: Hybrid AI Threats
Jeremy McHugh, Kristina Å ekrst, Jon Cefalu
Prompt injection attacks, where malicious input is designed to manipulate AI systems into ignoring their original instructions and following unauthorized commands instead, were fir…
cs.CY2024
AI Ethics by Design: Implementing Customizable Guardrails for Responsible AI Development
Kristina Å ekrst, Jeremy McHugh, Jonathan Rodriguez Cefalu
This paper explores the development of an ethical guardrail framework for AI systems, emphasizing the importance of customizable guardrails that align with diverse user values and…