1 paper · 1 filter
Yi Shen, Jessilyn Dunn, Michael M. Zavlanos
In this paper, we consider a risk-averse multi-armed bandit (MAB) problem where the goal is to learn a policy that minimizes the risk of low expected return, as opposed to maximizi…