2 papers
cs.SE2026
Grounding AI Agents in Contracts: An Empirical Evaluation of Spec-Driven Test Generation
Michele Tufano, James McClure, José Cambronero +7
LLM-based agents are increasingly used for coding tasks, where they have outperformed many classical approaches and scaled to repository-level tasks, such as test generation. Howev…
cs.SE2026
LLM-Based Automated Diagnosis Of Integration Test Failures At Google
Celal Ziftci, Ray Liu, Spencer Greene +1
Integration testing is critical for the quality and reliability of complex software systems. However, diagnosing their failures presents significant challenges due to the massive v…