1 paper · 1 filter
Huimin Xu, Xin Mao, Feng-Lin Li +4
Process Reward Models (PRMs) have demonstrated promising results in mathematical reasoning, but existing process annotation approaches, whether through human annotations or Monte C…