3 papers
cs.CR2026
Towards Reliable Local Security Agents: Verifiable Post-Training for Linux Privilege Escalation
Philipp Normann, Andreas Happe, Jürgen Cito +1
LLM agents are becoming increasingly important in the security domain, but leading systems are often closed-source, cloud-based, hard to reproduce or use with sensitive code. This…
cs.CR2026
Cochise: A Reference Harness for Autonomous Penetration Testing
Andreas Happe, Jürgen Cito, Jürgen Cito
Recent work on LLM-driven autonomous penetration testing reports promising results, but existing systems often bundle architectural, prompting, and tool-integration choices togethe…
cs.CR2026
Can LLMs Hack Enterprise Networks? -- Replicated Computational Results (RCR) Report
Andreas Happe, Jürgen Cito
This is the Replicated Computational Results (RCR) Report for the paper ``Can LLMs Hack Enterprise Networks?" The paper empirically investigates the efficacy and effectiveness of d…