3 citations · 3 across the 1 of their papers we have counts for
1 paper
Nobuhito Manome, Shuji Shinohara, Ung-il Chung
The multi-armed bandit (MAB) problem is a classical problem that models sequential decision-making under uncertainty in reinforcement learning. In this study, we propose a new gene…