7 citations · 8 across the 5 of their papers we have counts for
1 paper · 1 filter
Tong Mu, Georgios Theocharous, David Arbour +1
Online reinforcement learning (RL) algorithms are often difficult to deploy in complex human-facing applications as they may learn slowly and have poor early performance. To addres…