2 papers
cs.AI2026
When Skills Don't Help: A Negative Result on Procedural Knowledge for Tool-Grounded Agents in Offensive Cybersecurity
Samuel Jacob Chacko, James Hugglestone, Chashi Mahiul Islam +1
Agent Skills, structured packages of procedural knowledge loaded into an LLM agent at inference time, are widely reported to improve task pass rates by an average of 16.2~percentag…
cs.CR2026
STRIATUM-CTF: A Protocol-Driven Agentic Framework for General-Purpose CTF Solving
James Hugglestone, Samuel Jacob Chacko, Dawson Stoller +2
Large Language Models (LLMs) have demonstrated potential in code generation, yet they struggle with the multi-step, stateful reasoning required for offensive cybersecurity operatio…