2 papers
cs.LG2026
World-Time Compute with Verified Code World Models
James Schwoebel, Ingrida Semenec, Jenia Rousseva +8
LLMs generalize across a domain only after seeing many real, labeled examples, which most domains lack. We study a way to manufacture it cheaply. When a domain's dynamics can be wr…
cs.HC2025
From Prompt to Product: A Human-Centered Benchmark of Agentic App Generation Systems
Marcos Ortiz, Justin Hill, Collin Overbay +4
Agentic AI systems capable of generating full-stack web applications from natural language prompts ("prompt- to-app") represent a significant shift in software development. However…