Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Plan before Solving: Problem-Aware Strategy Routing for Mathematical Reasoning with LLMs
Shihao Qi, Jie Ma, Ziang Yin +5
Existing methods usually leverage a fixed strategy, such as natural language reasoning, code-augmented reasoning, tool-integrated reasoning, or ensemble-based reasoning, to guide L…
cs.AI2025
From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision
Jie Ma, Shihao Qi, Rui Xing +4
The quality of process data plays a key role in training a Process Reward Model (PRM), which can enhance the complex mathematical reasoning capability of large language models. Exi…