4 citations · 6 across the 2 of their papers we have counts for
4 papers · 1 filter
Offline RL With Resource Constrained Online Deployment
Jayanth Reddy Regatti, Aniket Anand Deshmukh, Frank Cheng +3
Offline reinforcement learning is used to train policies in scenarios where real-time access to the environment is expensive or impossible. As a natural consequence of these harsh…
Thompson Sampling in Non-Episodic Restless Bandits
Young Hun Jung, Marc Abeille, Ambuj Tewari
Restless bandit problems assume time-varying reward distributions of the arms, which adds flexibility to the model but makes the analysis more challenging. We study learning algori…
Regret Bounds for Thompson Sampling in Episodic Restless Bandit Problems
Young Hun Jung, Ambuj Tewari
Restless bandit problems are instances of non-stationary multi-armed bandits. These problems have been studied well from the optimization perspective, where the goal is to efficien…
Online Learning via the Differential Privacy Lens
Jacob Abernethy, Young Hun Jung, Chansoo Lee +2
In this paper, we use differential privacy as a lens to examine online learning in both full and partial information settings. The differential privacy framework is, at heart, less…