1 paper
Lexiang Tang, Weihao Gao, Bingchen Zhao +4
Recent work on test-time scaling for large language model (LLM) reasoning typically assumes that allocating more inference-time computation uniformly improves correctness. However,…