1 paper
Ziheng Qin, Yaxin Lu, Zhangyang Atlas Wang +1
Long-horizon agents can fail even when their underlying models can solve the constituent steps. They may lose track of mutable state, fail to reactivate lessons from earlier execut…