1 paper
Tom Perneczky, Marc Abeille, David Janz
We prove a variance-sensitive regret bound for Thompson sampling in stochastic generalised linear bandits. The argument assumes a warm-up, after which the regret is controlled thro…