1 paper · 1 filter
Matías Carrasco, Alejandro Cholaquidis
We study stochastic multi-armed bandits in which the objective is a statistical functional of the long-run reward distribution, rather than expected reward alone. Under mild contin…