1 paper · 1 filter
Nikolai Karpov, Chen Wang
We investigate the sample-memory-pass trade-offs for pure exploration in multi-pass streaming multi-armed bandits (MABs) with the *a priori* knowledge of the optimality gap $Î_{[2…