2 papers
cs.CY2026
AutoResearch: An Execution-Grounded Multi-Agent Framework for Reliable Research Workflow Automation
Rajesh Kumar, Waqar Ali, Junaid Ahmed +2
Automated research agents increasingly generate code, retrieve literature, and draft scientific artifacts, but they often fail to verify whether generated experiments execute corre…
cs.SE2026
AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering
Rajesh Kumar, Waqar Ali, Junaid Ahmed +2
Large language models generate plausible code but cannot verify correctness. Existing multi-agent systems simulate execution or leave verification optional. We introduce execution-…