3 papers
cs.AI2026
Intermediate Artifacts as First-Class Citizens: A Data Model for Durable Intermediate Artifacts in Agentic Systems
Josh Rosen, Seth Rosen
Many AI systems are organized around loops in which models reason, call tools, observe results, and continue until a task is complete. These systems often produce final artifacts s…
cs.AI2026
From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work
Josh Rosen, Seth Rosen
Large language model systems are increasingly deployed as agentic workflows that interleave reasoning, tool use, memory, and iterative refinement. These systems are effective at pr…
cs.CL2024
Seeing Through the Fog: A Cost-Effectiveness Analysis of Hallucination Detection Systems
Alexander Thomas, Seth Rosen, Vishnu Vettrivel
This paper presents a comparative analysis of hallucination detection systems for AI, focusing on automatic summarization and question answering tasks for Large Language Models (LL…