Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Scaffold Effects on GAIA: A Controlled Comparison
Jason Starace
Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well characterized under controll…
cs.AI2026
Intentional Deception as Controllable Capability in LLM Agents
Jason Starace, Terence Soule
As LLM-based agents increasingly operate in multi-agent systems, understanding adversarial manipulation becomes critical for defensive design. We present a systematic study of inte…