Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents
Qirui Mi, Zhijian Ma, Mengyue Yang +4
LLM-driven agents excel at sequential decision-making but often rely on on-the-fly reasoning, re-deriving solutions even in recurring scenarios. This insufficient experience reuse…
cs.AI2025
Learning to Discuss Strategically: A Case Study on One Night Ultimate Werewolf
Xuanfa Jin, Ziyan Wang, Yali Du +3
Communication is a fundamental aspect of human society, facilitating the exchange of information and beliefs among people. Despite the advancements in large language models (LLMs),…
cs.AI2024
Large Language Models Play StarCraft II: Benchmarks and A Chain of Summarization Approach
Weiyu Ma, Qirui Mi, Yongcheng Zeng +5
StarCraft II is a challenging benchmark for AI agents due to the necessity of both precise micro level operations and strategic macro awareness. Previous works, such as Alphastar a…