1 paper
Jie Ma, Shihao Qi, Rui Xing +4
The quality of process data plays a key role in training a Process Reward Model (PRM), which can enhance the complex mathematical reasoning capability of large language models. Exi…