1 paper
Priyal Deep, Shane Emmons, Amy Fox +4
LLM-powered applications routinely embed secrets in system prompts, yet models can be tricked into revealing them. We built an adaptive attacker that evolves its strategies over hu…