1 paper
Qirun Zeng, Eric He, Richard Hoffmann +2
Adversarial attacks on stochastic bandits have traditionally relied on some unrealistic assumptions, such as per-round reward manipulation and unbounded perturbations, limiting the…