2 papers
cs.SE2026
SWE-Test: Benchmarking LLM Vulnerability Discovery via Input Prediction
Yuanxiang Shi, Jiayi Lin, Xuanyong Lin +8
Vulnerability discovery is becoming an important ability of large language model (LLM) agents: agents that silently miss real defects leave critical software exposed. Rigorously me…
cs.SE2026
MOA: A Profiling-Guided LLM Framework for Memory-Optimization Automation at Codebase Scale
Jiaxi Liang, Yuanxiang Shi, Zezhou Yang +1
Modern large-scale software systems often suffer from pervasive memory inefficiencies (e.g., bloat, churn), leading to excessive resource costs and performance degradation. Existin…