1 paper
Zhen Yang, Mingyang Zhang, Feng Chen +4
Recent progress in large language models (LLMs) has focused on test-time scaling to improve reasoning via increased inference computation, but often at the cost of efficiency. We r…