15 citations · 28 across the 15 of their papers we have counts for
5 papers · 1 filter
Mars-PO: Multi-Agent Reasoning System Preference Optimization
Xiaoxuan Lou, Chaojie Wang, Bo An
Mathematical reasoning is a fundamental capability for large language models (LLMs), yet achieving high performance in this domain remains a significant challenge. The auto-regress…
In-Context Exploiter for Extensive-Form Games
Shuxin Li, Chang Yang, Youzhi Zhang +5
Nash equilibrium (NE) is a widely adopted solution concept in game theory due to its stability property. However, we observe that the NE strategy might not always yield the best re…
Grasper: A Generalist Pursuer for Pursuit-Evasion Problems
Pengdeng Li, Shuxin Li, Xinrun Wang +5
Pursuit-evasion games (PEGs) model interactions between a team of pursuers and an evader in graph-based environments such as urban street networks. Recent advancements have demonst…
Towards Skilled Population Curriculum for Multi-Agent Reinforcement Learning
Rundong Wang, Longtao Zheng, Wei Qiu +7
Recent advances in multi-agent reinforcement learning (MARL) allow agents to coordinate their behaviors in complex environments. However, common MARL algorithms still suffer from s…
Pretrained Cost Model for Distributed Constraint Optimization Problems
Yanchen Deng, Shufeng Kong, Bo An
Distributed Constraint Optimization Problems (DCOPs) are an important subclass of combinatorial optimization problems, where information and controls are distributed among multiple…