Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Can RL Improve Generalization of LLM Agents? An Empirical Study
Zhiheng Xi, Xin Guo, Jiaqi Liu +11
Reinforcement fine-tuning (RFT) has shown promise for training LLM agents to perform multi-turn decision-making based on environment feedback. However, most existing evaluations re…
cs.AI2025
MUA-RL: Multi-turn User-interacting Agent Reinforcement Learning for agentic tool use
Weikang Zhao, Xili Wang, Chengdi Ma +6
With the recent rapid advancement of Agentic Intelligence, agentic tool use in LLMs has become increasingly important. During multi-turn interactions between agents and users, the…