93 citations · 93 across the 1 of their papers we have counts for
1 paper
Amir Sani, Alessandro Lazaric, Rémi Munos
Stochastic multi-armed bandits solve the Exploration-Exploitation dilemma and ultimately maximize the expected reward. Nonetheless, in many practical problems, maximizing the expec…