1 paper
Eugene Lee, Oseong Choi, Byungsoo Kang +1
Multi-armed bandit algorithms, especially Thompson sampling, are widely used in online recommendation. Despite their ability to adapt from online feedback, these methods often suff…