44 citations · 127 across the 10 of their papers we have counts for
1 paper · 2 filters
Ronald Ortner, Daniil Ryabko, Peter Auer +1
We consider the restless Markov bandit problem, in which the state of each arm evolves according to a Markov process independently of the learner's actions. We suggest an algorithm…