8 citations · 9 across the 2 of their papers we have counts for
1 paper · 1 filter
Ruibo Yang, Jiazhou Wang, Andrew Mullhaupt
In this paper, we study the stochastic multi-armed bandit problem, where the reward is driven by an unknown random variable. We propose a new variant of the Upper Confidence Bound…