3 papers
cs.AI2026
Scaffold Effects on GAIA: A Controlled Comparison
Jason Starace
Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well characterized under controll…
cs.CY2026
Ethical Implications of Training Deceptive AI
Jason Starace, Bert Baumgaertner, Terence Soule
Deceptive behavior in AI systems is no longer theoretical: large language models strategically mislead without producing false statements, maintain deceptive strategies through saf…
cs.AI2026
Intentional Deception as Controllable Capability in LLM Agents
Jason Starace, Terence Soule
As LLM-based agents increasingly operate in multi-agent systems, understanding adversarial manipulation becomes critical for defensive design. We present a systematic study of inte…