Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents
Qirui Mi, Zhijian Ma, Mengyue Yang +4
LLM-driven agents excel at sequential decision-making but often rely on on-the-fly reasoning, re-deriving solutions even in recurring scenarios. This insufficient experience reuse…
cs.AI2026
Think, Speak, Decide: Language-Augmented Multi-Agent Reinforcement Learning for Economic Decision-Making
Heyang Ma, Qirui Mi, Qipeng Yang +3
Economic decision-making depends not only on structured signals such as prices and taxes, but also on unstructured language, including peer dialogue and media narratives. While mul…
cs.AI2024
Large Language Models Play StarCraft II: Benchmarks and A Chain of Summarization Approach
Weiyu Ma, Qirui Mi, Yongcheng Zeng +5
StarCraft II is a challenging benchmark for AI agents due to the necessity of both precise micro level operations and strategic macro awareness. Previous works, such as Alphastar a…