deterministic evaluation 1multi-agent systems 1orchestration planning 1simulation benchmark 1workflow scheduling 1
From the 1 of 6 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation
Zhenzhen Ren, Jiyan He, Xinpeng Zhang +5
The paper introduces OrchBench, a deterministic simulation benchmark that evaluates multi‑agent orchestration plans on DAG‑structured tasks in isolation, providing fast, token‑effi…
cs.AI2025
GTM: Simulating the World of Tools for AI Agents
Zhenzhen Ren, Xinpeng Zhang, Zhenxing Qian +4
The integration of external tools is pivotal for empowering Large Language Model (LLM) agents with real-world capabilities. However, training these agents through direct, continuou…