2 papers
cs.CL2026
DSPO: Stable and Efficient Policy Optimization for Agentic Search and Reasoning
Chenyang Gu, Yewen Pu, Bruce Yang +2
Enhancing LLMs with the ability to actively search external knowledge is crucial for complex and real-world tasks. Current approaches either rely on prompting to elicit the model's…
cs.AI2025
CodeAgents: A Token-Efficient Framework for Codified Multi-Agent Reasoning in LLMs
Bruce Yang, Xinfeng He, Huan Gao +3
Effective prompt design is essential for improving the planning capabilities of large language model (LLM)-driven agents. However, existing structured prompting strategies are typi…