46 citations
2 papers
cs.LG2020★ 46 cited
Whittle index based Q-learning for restless bandits with average reward
Konstantin E. Avrachenkov, Vivek S. Borkar
A novel reinforcement learning algorithm is introduced for multiarmed restless bandits with average reward, using the paradigms of Q-learning and Whittle index. Specifically, we le…
cs.SI2019★ 1 cited
How often should I access my online social networks?
Eduardo Hargreaves, Daniel Sadoc Menasché, Giovanni Neglia
Users of online social networks are faced with a conundrum of trying to be always informed without having enough time or attention budget to do so. The retention of users on online…