Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Bridging Rested and Restless Bandits with Graph-Triggering: Rising and Rotting
Gianmarco Genalti, Marco Mussi, Nicola Gatti +3
Rested and Restless Bandits are two well-known bandit settings that are useful to model real-world sequential decision-making problems in which the expected reward of an arm evolve…
stat.ML2025
Sliding-Window Thompson Sampling for Non-Stationary Settings
Marco Fiandri, Alberto Maria Metelli, Francesco Trovò
Non-stationary multi-armed bandits (NS-MABs) model sequential decision-making problems in which the expected rewards of a set of actions, a.k.a.~arms, evolve over time. In this pap…