Showing cs.AIShow all
2 papers · 1 filter
cs.AI2024
Mars-PO: Multi-Agent Reasoning System Preference Optimization
Xiaoxuan Lou, Chaojie Wang, Bo An
Mathematical reasoning is a fundamental capability for large language models (LLMs), yet achieving high performance in this domain remains a significant challenge. The auto-regress…
cs.AI2024
Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning
Chaojie Wang, Yanchen Deng, Zhiyi Lyu +4
Large Language Models (LLMs) have demonstrated impressive capability in many natural language tasks. However, the auto-regressive generation process makes LLMs prone to produce err…