10 citations · 10 across the 1 of their papers we have counts for
1 paper
Bin Chong, Yingguang Yang, Zi-Le Wang +2
Most algorithms for the multi-armed bandit problem in reinforcement learning aimed to maximize the expected reward, which are thus useful in searching the optimized candidate with…