1 paper · 1 filter
Yu-Han Huang, Argyrios Gerogiannis, Subhonmesh Bose +1
Conventional Multi-Armed Bandit (MAB) algorithms are designed for stationary environments, where the reward distributions associated with the arms do not change with time. In many…