1 paper
Yulian Wu, Chaowen Guan, Vaneet Aggarwal +1
In this paper, we study multi-armed bandits (MAB) and stochastic linear bandits (SLB) with heavy-tailed rewards and quantum reward oracle. Unlike the previous work on quantum bandi…