6 citations · 8 across the 6 of their papers we have counts for
Showing cs.SEShow all
2 papers · 1 filter
cs.SE2026★ 1 cited
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
Prithwish Jana, Sam Davidson, Bhavana Bhasker +3
Automating Infrastructure-as-Code (IaC) is challenging, and large language models (LLMs) often produce incorrect configurations from natural language (NL). We present TerraFormer,…
cs.SE2025★ 1 cited
SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
Muhammad Shihab Rashid, Christian Bock, Yuan Zhuang +10
Coding agents powered by large language models have shown impressive capabilities in software engineering tasks, but evaluating their performance across diverse programming languag…