1 paper · 1 filter
Anay Kulkarni, ChiaEn Lu, Dheeraj Mekala +3
Tool use enables large language models to solve complex tasks through sequences of API calls, yet existing reinforcement learning approaches fail to scale to multi-step composition…