1 paper
Kelly W. Zhang, Thomas Baldwin-McDonald, Kamil Ciosek +2
Increasingly, recommender systems are tasked with improving users' long-term satisfaction. In this context, we study a content exploration task, which we formalize as a bandit prob…