11 citations · 14 across the 3 of their papers we have counts for
1 paper · 1 filter
Rémy Degenne, Han Shao, Wouter M. Koolen
We study reward maximisation in a wide class of structured stochastic multi-armed bandit problems, where the mean rewards of arms satisfy some given structural constraints, e.g. li…