1 paper
Youness Bouchari, Matteo Boffa, Marco Mellia +3
Large Language Model (LLM) agents are increasingly proposed to automate offensive security tasks, with recent studies reporting near human-level success rates in Capture-the-Flag (…