52 citations · 101 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
GEAR: A GPU-Centric Experience Replay System for Large Reinforcement Learning Models
Hanjing Wang, Man-Kit Sit, Congjie He +5
This paper introduces a distributed, GPU-centric experience replay system, GEAR, designed to perform scalable reinforcement learning (RL) with large sequence models (such as transf…
cs.LG2016★ 52 cited
Product-based Neural Networks for User Response Prediction
Yanru Qu, Han Cai, Kan Ren +4
Predicting user responses, such as clicks and conversions, is of great importance and has found its usage in many Web applications including recommender systems, web search and onl…