2 papers
cs.SE2026
SWE-InfraBench: Evaluating Language Models on Cloud Infrastructure Code
Natalia Tarasova, Enrique Balp-Straffon, Aleksei Iancheruk +10
Building infrastructure-as-code (IaC) in cloud computing is a critical task, underpinning the reliability, scalability, and security of modern software systems. Despite the remarka…
cs.LG2025
Enabling Fine-Grained Operating Points for Black-Box LLMs
Ege Beyazit, KL Navaneet, Prashant Mathur +3
Black-box Large Language Models (LLMs) provide practical and accessible alternatives to other machine learning methods, as they require minimal labeled data and machine learning ex…