1 paper
Deqi Zheng, Xiaoyang Xu, Yuhong Yang
We study a stochastic multi-armed bandit problem in which the set of available arms expands over time. This setting arises in sequential experimentation when new actions or treatme…