Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
MCP-Universe RL: A Framework for Training MCP Tool-Use Agents via Reinforcement Learning
Ziyang Luo, Yan Yang, Xiangru Jian +5
Reinforcement learning (RL) has become an effective way to improve the tool-use ability of large language models (LLMs), but most existing RL frameworks stop at the policy update.…
cs.AI2026
W&D:Scaling Parallel Tool Calling for Efficient Deep Research Agents
Xiaoqiang Lin, Jun Hao Liew, Silvio Savarese +1
Deep research agents have emerged as powerful tools for automating complex intellectual tasks through multi-step reasoning and web-based information seeking. While recent efforts h…