1 paper
Ali Al-Kaswan, Maksim Plotnikov, Maxim Hájek +3
Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings remain poorly understood. We…