5 citations · 9 across the 4 of their papers we have counts for
3 papers · 1 filter
Online Model Selection: a Rested Bandit Formulation
Leonardo Cella, Claudio Gentile, Massimiliano Pontil
Motivated by a natural problem in online model selection with bandit information, we introduce and analyze a best arm identification problem in the rested bandit setting, wherein a…
Meta-learning with Stochastic Linear Bandits
Leonardo Cella, Alessandro Lazaric, Massimiliano Pontil
We investigate meta-learning procedures in the setting of stochastic linear bandits tasks. The goal is to select a learning algorithm which works well on average over a class of ba…
Stochastic Bandits with Delay-Dependent Payoffs
Leonardo Cella, Nicolò Cesa-Bianchi
Motivated by recommendation problems in music streaming platforms, we propose a nonstationary stochastic bandit model in which the expected reward of an arm depends on the number o…