2 papers
cs.LG2026
GEPO: Group Expectation Policy Optimization for Stable Heterogeneous Reinforcement Learning
Han Zhang, Ruibin Zheng, Zexuan Yi +16
As single-center computing approaches power constraints, decentralized training becomes essential. However, traditional Reinforcement Learning (RL) methods, crucial for enhancing l…
cs.DC2025
Static Batching of Irregular Workloads on GPUs: Framework and Application to Efficient MoE Model Inference
Yinghan Li, Yifei Li, Jiejing Zhang +13
It has long been a problem to arrange and execute irregular workloads on massively parallel devices. We propose a general framework for statically batching irregular workloads into…