1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CR2025
Prompt Injection 2.0: Hybrid AI Threats
Jeremy McHugh, Kristina Šekrst, Jon Cefalu
Prompt injection attacks, where malicious input is designed to manipulate AI systems into ignoring their original instructions and following unauthorized commands instead, were fir…
cs.CY2024★ 1 cited
AI Ethics by Design: Implementing Customizable Guardrails for Responsible AI Development
Kristina Šekrst, Jeremy McHugh, Jonathan Rodriguez Cefalu
This paper explores the development of an ethical guardrail framework for AI systems, emphasizing the importance of customizable guardrails that align with diverse user values and…