3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.LG2021
Batched Thompson Sampling
Cem Kalkanli, Ayfer Ozgur
We introduce a novel anytime Batched Thompson sampling policy for multi-armed bandits where the agent observes the rewards of her actions and adjusts her policy only at the end of…
cs.LG2021★ 1 cited
Asymptotic Performance of Thompson Sampling in the Batched Multi-Armed Bandits
Cem Kalkanli, Ayfer Ozgur
We study the asymptotic performance of the Thompson sampling algorithm in the batched multi-armed bandit setting where the time horizon is divided into batches, and the agent i…
cs.LG2020★ 3 cited
Asymptotic Convergence of Thompson Sampling
Cem Kalkanli, Ayfer Ozgur
Thompson sampling has been shown to be an effective policy across a variety of online learning tasks. Many works have analyzed the finite time performance of Thompson sampling, and…