1 paper
Yufan Wu, Yinghui He, Zhengyi Hu +4
Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated g…