3 papers
cs.AI2026
Agentic Time Machine as an Infrastructure for Future-Event Forecasting
Jingyi Chai, Bingyang Zheng, Xiangrui Liu +5
Forecasting future events is a critical challenge for large language model (LLM) agents, spanning domains from elections and monetary policy to financial markets. However, evaluati…
cs.AI2026
World of Workflows: A Benchmark for Bringing World Models to Enterprise Systems
Lakshya Gupta, Litao Li, Yizhe Liu +5
Frontier large language models (LLMs) excel as autonomous agents in many domains, yet they remain untested in complex enterprise systems where hidden workflows create cascading eff…
cs.AI2025
SCOPE: Language Models as One-Time Teacher for Hierarchical Planning in Text Environments
Haoye Lu, Pavan Seshadri, Kaheer Suleman
Long-term planning in complex, text-based environments presents significant challenges due to open-ended action spaces, ambiguous observations, and sparse feedback. Recent research…