deterministic evaluation 1multi-agent systems 1orchestration planning 1simulation benchmark 1workflow scheduling 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation
Zhenzhen Ren, Jiyan He, Xinpeng Zhang +5
The paper introduces OrchBench, a deterministic simulation benchmark that evaluates multi‑agent orchestration plans on DAG‑structured tasks in isolation, providing fast, token‑effi…
cs.CR2025
CoTSRF: Utilize Chain of Thought as Stealthy and Robust Fingerprint of Large Language Models
Zhenzhen Ren, GuoBiao Li, Sheng Li +2
Despite providing superior performance, open-source large language models (LLMs) are vulnerable to abusive usage. To address this issue, recent works propose LLM fingerprinting met…
cs.CV2025
Adversarial Shallow Watermarking
Guobiao Li, Lei Tan, Yuliang Xue +4
Recent advances in digital watermarking make use of deep neural networks for message embedding and extraction. They typically follow the ``encoder-noise layer-decoder''-based archi…