Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?
Dong Yan, Jian Liang, Dapeng Hu +4
Large language model (LLM) agents can self-evolve by continually improving from their own accumulated experience. However, existing studies predominantly adopt independent evaluati…
cs.AI2026
WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis
Shuo Lu, Yinuo Xu, Kecheng Yu +8
Large language models (LLMs) are increasingly asked not only to write static interfaces, but to construct executable interactive worlds from natural language. Browser-native 3D, co…