1 paper · 1 filter
Zijing Zhang, Ziyang Chen, Mingxiao Li +2
The development of autonomous agents for complex, long-horizon tasks is a central goal in AI. However, dominant training paradigms face a critical limitation: reinforcement learnin…