178 citations · 428 across the 16 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2023
Fast Rates for Maximum Entropy Exploration
Daniil Tiapkin, Denis Belomestny, Daniele Calandriello +7
We address the challenge of exploration in reinforcement learning (RL) when the agent operates in an unknown environment with sparse or no rewards. In this work, we study the maxim…
stat.ML2012★ 36 cited
Thompson Sampling: An Asymptotically Optimal Finite Time Analysis
Emilie Kaufmann, Nathaniel Korda, Rémi Munos
The question of the optimality of Thompson Sampling for solving the stochastic multi-armed bandit problem had been open since 1933. In this paper we answer it positively for the ca…