2 papers
cs.LG2021
Distributed Online Learning for Joint Regret with Communication Constraints
Dirk van der Hoeven, Hédi Hadiji, Tim van Erven
We consider distributed online learning for joint regret with communication constraints. In this setting, there are multiple agents that are connected in a graph. Each round, an ad…
stat.ML2019
Polynomial Cost of Adaptation for X -Armed Bandits
Hédi Hadiji
In the context of stochastic continuum-armed bandits, we present an algorithm that adapts to the unknown smoothness of the objective function. We exhibit and compute a polynomial c…