1 citations · 1 across the 5 of their papers we have counts for
1 paper · 1 filter
Yuxiang Ji, Ziyu Ma, Yong Wang +3
Recent advances in reinforcement learning (RL) have significantly enhanced the agentic capabilities of large language models (LLMs). In long-term and multi-turn agent tasks, existi…