Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Grounded Scaling: Why Agentic AI Needs Deterministic Environments
Liang Ding, Xintong Wang
Long-chain agent execution fails exponentially in environments designed for human tolerance: with per-step determinism , -step chain success degrades as . The AGI-t…
cs.AI2026
IndustryBench: Probing the Industrial Knowledge Boundaries of LLMs
Songlin Bai, Xintong Wang, Linlin Yu +12
In industrial procurement, an LLM answer is useful only if it survives a standards check: recommended material must match operating condition, every parameter must respect a regula…