21 citations · 43 across the 20 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
cs.LG2018
Tsallis-INF: An Optimal Algorithm for Stochastic and Adversarial Bandits
Julian Zimmert, Yevgeny Seldin
We derive an algorithm that achieves the optimal (within constants) pseudo-regret in both adversarial and stochastic multi-armed bandits without prior knowledge of the regime and t…
cs.LG2018
Factored Bandits
Julian Zimmert, Yevgeny Seldin
We introduce the factored bandits model, which is a framework for learning with limited (bandit) feedback, where actions can be decomposed into a Cartesian product of atomic action…