From the 1 of 17 linked papers with an AI index.
Showing cs.SEShow all
2 papers · 1 filter
cs.SE2026
RepoMod-Bench: A Benchmark for Code Repository Modernization via Implementation-Agnostic Testing
Xuefeng Li, Nir Ben-Israel, Yotam Raz +3
The evolution of AI coding agents has shifted the frontier from simple snippet completion to autonomous repository-level engineering. However, evaluating these agents remains ill-p…
cs.SE2026
daVinci-Dev: Agent-native Mid-training for Software Engineering
Ji Zeng, Dayuan Fu, Tiantian Mi +14
Recently, the frontier of Large Language Model (LLM) capabilities has shifted from single-turn code generation to agentic software engineering-a paradigm where models autonomously…