1 paper
Zicheng Liu, Lige Huang, Jie Zhang +3
The increasing autonomy of Large Language Models (LLMs) necessitates a rigorous evaluation of their potential to aid in cyber offense. Existing benchmarks often lack real-world com…