1 paper
Shogo Ito, Tatsuji Takahashi, Yu Kono
The contextual bandit problem, which is a type of reinforcement learning tasks, provides an effective framework for solving challenges in recommendation systems, such as satisfying…