2 papers
cs.CR2025
Prompt Injection 2.0: Hybrid AI Threats
Jeremy McHugh, Kristina Å ekrst, Jon Cefalu
Prompt injection attacks, where malicious input is designed to manipulate AI systems into ignoring their original instructions and following unauthorized commands instead, were fir…
cs.CY2024
AI Ethics by Design: Implementing Customizable Guardrails for Responsible AI Development
Kristina Å ekrst, Jeremy McHugh, Jonathan Rodriguez Cefalu
This paper explores the development of an ethical guardrail framework for AI systems, emphasizing the importance of customizable guardrails that align with diverse user values and…