1 paper
Xinjie Shen, Wei Fan, Xudong Guo +5
Language-model agents increasingly face long-horizon tasks with evolving state, interdependent decisions, and delayed outcomes. Scaling their training requires diverse agentic envi…