1 paper
Qiao Liang, Yuke Zhu, Chao Ge +4
Tool-integrated reasoning (TIR) enables LLM agents to solve tasks through planning, tool use, and iterative revision, but outcome-only reinforcement learning in this setting suffer…