1 paper · 1 filter
Anshuka Rangi, Massimo Franceschetti, Long Tran-Thanh
In this work, we study sequential choice bandits with feedback. We propose bandit algorithms for a platform that personalizes users' experience to maximize its rewards. For each ac…