1 paper
Zhongyi Li, Wan Tian, Jinju Chen +4
Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such systems is limited by credit assi…