1 paper
Yushu Li, Wenlong Deng, Jiajin Li +1
Test-time scaling has become a dominant paradigm for improving LLM agent reliability, yet current approaches treat compute as an abundant resource, allowing agents to exhaust token…