3 citations · 3 across the 1 of their papers we have counts for
2 papers
cs.LG2019
Regret Minimisation in Multi-Armed Bandits Using Bounded Arm Memory
Arghya Roy Chaudhuri, Shivaram Kalyanakrishnan
In this paper, we propose a constant word (RAM model) algorithm for regret minimisation for both finite and infinite Stochastic Multi-Armed Bandit (MAB) instances. Most of the exis…
cs.LG2019★ 3 cited
PAC Identification of Many Good Arms in Stochastic Multi-Armed Bandits
Arghya Roy Chaudhuri, Shivaram Kalyanakrishnan
We consider the problem of identifying any out of the best arms in an -armed stochastic multi-armed bandit. Framed in the PAC setting, this particular problem generalise…