6 citations · 6 across the 1 of their papers we have counts for
1 paper
Jonathan Cuevas, Ryugo Iwami, Atsushi Uchida +2
The Multi-Armed Bandit (MAB) problem, foundational to reinforcement learning-based decision-making, addresses the challenge of maximizing rewards amidst multiple uncertain choices.…