1 citations · 1 across the 1 of their papers we have counts for
1 paper
Jing Dong, Jingwei Li, Baoxiang Wang +1
Reinforcement learning (RL) has exceeded human performance in many synthetic settings such as video games and Go. However, real-world deployment of end-to-end RL models is less com…