3 citations · 6 across the 3 of their papers we have counts for
3 papers
AutoPenBench: Benchmarking Generative Agents for Penetration Testing
Luca Gioacchini, Marco Mellia, Idilio Drago +3
Generative AI agents, software systems powered by Large Language Models (LLMs), are emerging as a promising approach to automate cybersecurity tasks. Among the others, penetration…
AgentQuest: A Modular Benchmark Framework to Measure Progress and Improve LLM Agents
Luca Gioacchini, Giuseppe Siracusano, Davide Sanvito +4
The advances made by Large Language Models (LLMs) have led to the pursuit of LLM agents that can solve intricate, multi-step reasoning tasks. As with any research pursuit, benchmar…
Benchmarking Evolutionary Community Detection Algorithms in Dynamic Networks
Giordano Paoletti, Luca Gioacchini, Marco Mellia +2
In dynamic complex networks, entities interact and form network communities that evolve over time. Among the many static Community Detection (CD) solutions, the modularity-based Lo…