1 paper · 1 filter
Shubhada Agrawal, Timothée Mathieu, Debabrota Basu +1
We investigate the regret-minimisation problem in a multi-armed bandit setting with arbitrary corruptions. Similar to the classical setup, the agent receives rewards generated inde…