3 papers
cs.LG2026
Stagnant Neuron: Towards Understanding the Plasticity Loss in Multi-Agent Reinforcement Learning Value Factorization Methods
Zhengzhu Liu, Zeming Gao, Haoyuan Qin +7
Multi-Agent Reinforcement Learning (MARL) value factorization methods can suffer from a loss of plasticity, gradually failing to adapt when transferring to new task instances. We t…
cs.LG2026
MAGE: Multi-scale Autoregressive Generation for Offline Reinforcement Learning
Chenxing Lin, Xinhui Gao, Haipeng Zhang +7
Generative models have gained significant traction in offline reinforcement learning (RL) due to their ability to model complex trajectory distributions. However, existing generati…
cs.AI2025
PlanU: Large Language Model Reasoning through Planning under Uncertainty
Ziwei Deng, Mian Deng, Chenjing Liang +7
Large Language Models (LLMs) are increasingly being explored across a range of reasoning tasks. However, LLMs sometimes struggle with reasoning tasks under uncertainty that are rel…