86 citations · 101 across the 6 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
cs.LG2022★ 14 cited
EnvPool: A Highly Parallel Reinforcement Learning Environment Execution Engine
Jiayi Weng, Min Lin, Shengyi Huang +9
There has been significant progress in developing reinforcement learning (RL) training systems. Past works such as IMPALA, Apex, Seed RL, Sample Factory, and others, aim to improve…
cs.LG2022
A2C is a special case of PPO
Shengyi Huang, Anssi Kanervisto, Antonin Raffin +3
Advantage Actor-critic (A2C) and Proximal Policy Optimization (PPO) are popular deep reinforcement learning algorithms used for game AI in recent years. A common understanding is t…