From the 1 of 10 linked papers with an AI index.
10 papers
Hierarchical Graph Memory for LLM Agents with Path-level Localization and Rewrite
Xiawei Yue, Boran Wang, Xiaoqing Zhang +2
Agents for long term reasoning require a memory that can be efficiently and effectively updated over time, as new facts and external feedback continue to arrive. Recently, graph me…
OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation
Zhenzhen Ren, Jiyan He, Xinpeng Zhang +5
The paper introduces OrchBench, a deterministic simulation benchmark that evaluates multi‑agent orchestration plans on DAG‑structured tasks in isolation, providing fast, token‑effi…
Modeling Earth-Scale Human-Like Societies with One Billion Agents
Haoxiang Guan, Jiyan He, Liyang Fan +10
Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulations. Traditional agent-based models (…
APS: Bias-Controlled Adaptive Prototype Simulation for Population-Scale LLM Agents
Quan Zheng, Yan Gao, Shaobin He +6
LLM-agent simulation offers a flexible computational tool for studying population response trajectories that depend on scenario events, memory, demographics, and evolving social co…
FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards
Zhixin Han, Yanzhi Zhang, Chuyang Wei +11
Live future prediction refers to the task of making predictions about real-world events before they unfold. This task is increasingly studied using large language model-based agent…
GUIGuard-Bench: Toward a General Evaluation for Privacy-Preserving GUI Agents
Yanxi Wang, Zhiling Zhang, Wenbo Zhou +6
As GUI agents increasingly rely on screenshots to perceive and operate digital environments, they may inadvertently expose sensitive information such as identities, accounts, locat…