178 citations · 427 across the 8 of their papers we have counts for
1 paper · 1 filter
Amir Sani, Alessandro Lazaric, Rémi Munos
Stochastic multi-armed bandits solve the Exploration-Exploitation dilemma and ultimately maximize the expected reward. Nonetheless, in many practical problems, maximizing the expec…