Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
Backdoor Decontamination Dynamics in LLM Agents
Gabriel Huang, Abhay Puri, Léo Boisvert +4
Open-weight LLM agents are vulnerable to backdoors installed during fine-tuning, which may be undetectable if the trigger conditions are never met during testing. Assuming defender…
cs.CR2025
DoomArena: A framework for Testing AI Agents Against Evolving Security Threats
Leo Boisvert, Mihir Bansal, Chandra Kiran Reddy Evuru +9
We present DoomArena, a security evaluation framework for AI agents. DoomArena is designed on three principles: 1) It is a plug-in framework and integrates easily into realistic ag…