1 paper
Apurva Gandhi, Satyaki Chakraborty, Xiangjun Wang +2
We introduce Recursive Agent Optimization (RAO), a reinforcement learning approach for training recursive agents: agents that can spawn and delegate sub-tasks to new instantiations…