1 paper
Kaiyuan Liu, Qiuyang Mang, Bo Peng +6
Large language model (LLM) agents allocate test-time compute adaptively as they revise solutions, use tools, explore alternatives, and decide when to stop. This test-time strategy…