Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Symphony: A Decentralized Multi-Agent Framework for Scalable Collective Intelligence
Ji Wang, Kashing Chen, Xinyuan Song +4
Most existing Large Language Model (LLM)-based agent frameworks rely on centralized orchestration, incurring high deployment costs, rigid communication topologies, and limited adap…
cs.LG2025
Echo: Decoupling Inference and Training for Large-Scale RL Alignment on Heterogeneous Swarms
Jie Xiao, Changyuan Fan, Qingnan Ren +6
Modern RL-based post-training for large language models (LLMs) co-locate trajectory sampling and policy optimisation on the same GPU cluster, forcing the system to switch between i…