1 paper · 1 filter
Yufan Wu, Yinghui He, Zhengyi Hu +4
Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated g…