Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
The World Won't Stay Still: Programmable Evolution for Agent Benchmarks
Guangrui Li, Yaochen Xie, Yi Liu +11
LLM-powered tool-calling agents fulfill user requests by interacting with environments, querying data, and invoking tools in a multi-turn process. Yet, most existing benchmarks eva…
cs.AI2025
Scaling Agent Learning via Experience Synthesis
Zhaorun Chen, Zhuokai Zhao, Kai Zhang +15
While reinforcement learning (RL) can empower autonomous agents by enabling self-improvement through interaction, its practical adoption remains challenging due to costly rollouts,…