1 paper
Keqin Liu, Tianshuo Zheng, Zhi-Hua Zhou
The multi-armed bandit (MAB) problems are widely studied in fields of operations research, stochastic optimization, and reinforcement learning. In this paper, we consider the class…