2 papers
cs.LG2026
Rollout-Training Co-Design for Efficient LLM-Based Multi-Agent Reinforcement Learning
Zhida Jiang, Zhaolong Xing, Jiawei Lu +13
Despite algorithm-level innovations for multi-agent reinforcement learning (MARL), the underlying networked infrastructure for large-scale MARL training remains underexplored. Exis…
quant-ph2026
Learning to Decode in Parallel: Self-Coordinating Neural Network for Real-Time Quantum Error Correction
Kai Zhang, Zhengzhong Yi, Shaojun Guo +13
Fast, reliable decoders are pivotal components for enabling fault-tolerant quantum computation (FTQC). Neural network decoders like AlphaQubit have demonstrated potential, achievin…