Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning
Xiaoxuan Wang, Han Zhang, Haixin Wang +11
Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interactive tasks. Despite encouraging ea…
cs.AI2026
FitText: Evolving Agent Tool Ecologies via Memetic Retrieval
Kyle Zheng, Han Zhang, Renliang Sun +2
Efficient reasoning is not only a matter of shortening an answer trace; for tool-using agents, it also depends on whether the agent is reasoning over the right action space. As API…