5 citations · 5 across the 2 of their papers we have counts for
1 paper · 1 filter
Hanjing Wang, Man-Kit Sit, Congjie He +5
This paper introduces a distributed, GPU-centric experience replay system, GEAR, designed to perform scalable reinforcement learning (RL) with large sequence models (such as transf…