1 paper
Aleksandr Vorobev, Gleb Gusev
We study the stochastic multi-armed bandit problem with non-equivalent multiple plays where, at each step, an agent chooses not only a set of arms, but also their order, which infl…