1 paper · 1 filter
Yunpeng Dong, Jingkai He, Shiqi Liu +7
LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid checkpoint and rollback (C/R) of the compl…