1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.MA2026
Adaptive Value Decomposition: Coordinating a Varying Number of Agents in Urban Systems
Yexin Li, Jinjin Guo, Haoyu Zhang +3
Multi-agent reinforcement learning (MARL) provides a promising paradigm for coordinating multi-agent systems (MAS). However, most existing methods rely on restrictive assumptions,…
cs.AI2025
AutoPBO: LLM-powered Optimization for Local Search PBO Solvers
Jinyuan Li, Yi Chu, Yiwen Sun +2
Pseudo-Boolean Optimization (PBO) provides a powerful framework for modeling combinatorial problems through pseudo-Boolean (PB) constraints. Local search solvers have shown excelle…
cs.AI2024★ 1 cited
Improving Multi-Step Reasoning Abilities of Large Language Models with Direct Advantage Policy Optimization
Jiacai Liu, Chaojie Wang, Chris Yuhao Liu +5
The role of reinforcement learning (RL) in enhancing the reasoning of large language models (LLMs) is becoming increasingly significant. Despite the success of RL in many scenarios…