activity
20232026
most citedControlling risks of AI in chemical science with agents

6 citations · 6 across the 13 of their papers we have counts for

collaborators

13 papers

cs.AI2026

AgentRewind: Recoverable Execution for Long-Horizon LLM Agents

Yu Zhuang, Kefei Chen, Yitong Duan +3

Many real-world tasks require LLM agents to interact with their environments over long execution horizons. Errors that occur early in execution may propagate through both the agent…

cs.AI2026

Hierarchical Graph Memory for LLM Agents with Path-level Localization and Rewrite

Xiawei Yue, Boran Wang, Xiaoqing Zhang +2

Agents for long term reasoning require a memory that can be efficiently and effectively updated over time, as new facts and external feedback continue to arrive. Recently, graph me…

cs.AI2026

OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation

Zhenzhen Ren, Jiyan He, Xinpeng Zhang +5

Complex tasks often decompose into parallelizable yet interdependent subtasks, making orchestration critical to the performance of multi-agent systems (MAS). Existing evaluations t…

cs.MA2026

APS: Bias-Controlled Adaptive Prototype Simulation for Population-Scale LLM Agents

Quan Zheng, Yan Gao, Shaobin He +6

LLM-agent simulation offers a flexible computational tool for studying population response trajectories that depend on scenario events, memory, demographics, and evolving social co…

cs.AI2026

FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards

Zhixin Han, Yanzhi Zhang, Chuyang Wei +11

Live future prediction refers to the task of making predictions about real-world events before they unfold. This task is increasingly studied using large language model-based agent…

cs.AI2026

Harnessing Pre-Resolution Signals for Future Prediction Agents

Chuyang Wei, Maohang Gao, Zhixin Han +12

Many high-stakes decisions depend on forecasts made before outcomes are known. In this future prediction setting, the central challenge is that public evidence evolves over time, w…