1 paper · 1 filter
Yingchuan Zhang, Terry Ma, Wenxuan Zhong +1
Large language models (LLMs) achieve higher accuracy on challenging reasoning tasks by scaling test-time compute through multiple trajectory sampling. However, standard aggregation…