1 paper · 1 filter
Simon Buchholz, Jonas M. Kübler, Bernhard Schölkopf
Multi-armed bandits are one of the theoretical pillars of reinforcement learning. Recently, the investigation of quantum algorithms for multi-armed bandit problems was started, and…