Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Potentially Optimal Joint Actions Recognition for Cooperative Multi-Agent Reinforcement Learning
Chang Huang, Shatong Zhu, Junqiao Zhao +6
Value function factorization is widely used in cooperative multi-agent reinforcement learning (MARL). Existing approaches often impose monotonicity constraints between the joint ac…
cs.LG2026
ACSAC: Adaptive Chunk Size Actor-Critic with Causal Transformer Q-Network
Qian Chen, Junqiao Zhao, Hongtu Zhou +4
Long-horizon, sparse-reward tasks pose a fundamental challenge for reinforcement learning, since single-step TD learning suffers from bootstrapping error accumulation across succes…