3 papers
cs.RO2026
IndoorR2X: Indoor Robot-to-Everything Coordination with LLM-Driven Planning
Fan Yang, Soumya Teotia, Shaunak A. Mehta +8
Although robot-to-robot (R2R) communication improves indoor scene understanding beyond what a single robot can achieve, R2R alone cannot overcome partial observability without subs…
cs.AI2025
Plan before Solving: Problem-Aware Strategy Routing for Mathematical Reasoning with LLMs
Shihao Qi, Jie Ma, Ziang Yin +5
Existing methods usually leverage a fixed strategy, such as natural language reasoning, code-augmented reasoning, tool-integrated reasoning, or ensemble-based reasoning, to guide L…
cs.AI2025
From Static to Dynamic: Adaptive Monte Carlo Search for Mathematical Process Supervision
Jie Ma, Shihao Qi, Rui Xing +4
The quality of process data plays a key role in training a Process Reward Model (PRM), which can enhance the complex mathematical reasoning capability of large language models. Exi…