3 papers
cs.AI2026
The Greatness of Science Cannot Be Planned: Agentic Auto-Research is Fuzz Testing
Yifeng He, Jicheng Wang, Yinzhe Zhao +3
Agentic auto-research is emerging, but most systems treat scientific discovery as goal-oriented optimization against a final benchmark. This paradigm rewards a sparse final verdict…
cs.SE2026
VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection
Kexing Ji, Jiachen Liu, Enze Hu +7
Recent advances in LLM-based vulnerability detection have shown promising results, while coding agents further extend this capability from isolated code snippets to complete reposi…
cs.SE2026
Towards Reliable C-to-Rust Translation with Rule-Guided Reasoning and Reinforcement Learning
Feng Luo, Jiachen Liu, Cuiyun Gao +2
The migration of legacy C programs to Rust has become an important direction for improving software memory safety while alleviating the high cost of manual rewriting. Leveraging la…