3 papers
cs.AI2026
Efficiency Matters in Autonomous Research
Haiqian Yang, Yuan Cao
AI-driven autonomous research (AR) systems are becoming increasingly effective across a broad range of tasks. Their performance, however, is still evaluated primarily by the qualit…
cs.AI2026
Beyond Fixed Representations: The Vocabulary and Verifier Gaps in Open-Ended AI
Yuan Cao, Haiqian Yang
Modern AI systems are increasingly being evaluated for their ability to reason, code, prove theorems, use tools, and long-horizon research tasks. These are powerful capabilities, b…
cs.SE2026
A Survey of LLM-Driven Penetration Testing: Taxonomy, Co-Evolution, and Open Challenges
Zheyuan He, Jiaxun Dong, Zihao Li +6
Agents4Pentest, an emerging class of LLM-based autonomous penetration testing systems, has become a rapidly growing area in security research. Despite this growth, the field still…