3 papers
cs.MA2026
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
Yuanjun Li, Bin Zhang, Hao Chen +3
Value decomposition (VD) methods have achieved remarkable success in cooperative multi-agent reinforcement learning (MARL). However, their reliance on the max operator for temporal…
cs.CL2025
Balancing Rewards in Text Summarization: Multi-Objective Reinforcement Learning via HyperVolume Optimization
Junjie Song, Yiwen Liu, Dapeng Li +4
Text summarization is a crucial task that requires the simultaneous optimization of multiple objectives, including consistency, coherence, relevance, and fluency, which presents co…
cs.AI2025
Adaptive parameter sharing for multi-agent reinforcement learning
Dapeng Li, Na Lou, Bin Zhang +2
Parameter sharing, as an important technique in multi-agent systems, can effectively solve the scalability issue in large-scale agent problems. However, the effectiveness of parame…