1 paper
Sourav Chakraborty, Lijun Chen
We study incentivized exploration for the multi-armed bandit (MAB) problem with non-stationary reward distributions, where players receive compensation for exploring arms other tha…