2 papers
cs.SE2026
A Multi-Agent Framework for Automated Exploit Generation with Constraint-Guided Comprehension and Reflection
Siyi Chen, Tianhan Luo, Shijian Wu +4
Open-source libraries are widely used in modern software development, introducing significant security vulnerabilities. While static analysis tools can identify potential vulnerabi…
cs.CR2026
SecCodeBench-V2 Technical Report
Longfei Chen, Ji Zhao, Lanxiao Cui +24
We introduce SecCodeBench-V2, a publicly released benchmark for evaluating Large Language Model (LLM) copilots' capabilities of generating secure code. SecCodeBench-V2 comprises 98…