Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Recursive Agent Optimization
Apurva Gandhi, Satyaki Chakraborty, Xiangjun Wang +2
We introduce Recursive Agent Optimization (RAO), a reinforcement learning approach for training recursive agents: agents that can spawn and delegate sub-tasks to new instantiations…
cs.LG2025
Shared DIFF Transformer
Yueyang Cang, Yuhang Liu, Xiaoteng Zhang +2
DIFF Transformer improves attention allocation by enhancing focus on relevant context while suppressing noise. It introduces a differential attention mechanism that calculates the…