2 papers
cs.CR2026
AgentCyberRange: Benchmarking Frontier AI Systems in Realistic Cyber Ranges
Fengyu Liu, Jiarun Dai, Yihe Fan +11
Frontier AI systems are increasingly capable of cybersecurity tasks, including codebase inspection, vulnerability detection, and exploitation. However, evaluating their offensive c…
cs.CR2024
When LLMs Meet Cybersecurity: A Systematic Literature Review
Jie Zhang, Haoyu Bu, Hui Wen +7
The rapid development of large language models (LLMs) has opened new avenues across various fields, including cybersecurity, which faces an evolving threat landscape and demand for…