44 citations · 155 across the 13 of their papers we have counts for
1 paper · 1 filter
Ronald Ortner, Daniil Ryabko, Peter Auer +1
We consider the restless Markov bandit problem, in which the state of each arm evolves according to a Markov process independently of the learner's actions. We suggest an algorithm…