6 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Lagrangian Index Policy for Restless Bandits with Average Reward
Konstantin Avrachenkov, Vivek S. Borkar, Pratik Shah
We study the Lagrangian Index Policy (LIP) for restless multi-armed bandits with long-run average reward. In particular, we compare the performance of LIP with the performance of t…
cs.AI2024★ 6 cited
Tabular and Deep Learning for the Whittle Index
Francisco Robledo Relaño, Vivek Borkar, Urtzi Ayesta +1
The Whittle index policy is a heuristic that has shown remarkably good performance (with guaranteed asymptotic optimality) when applied to the class of problems known as Restless M…