1 paper
Sushant Vijayan, Arun Suggala, Karthikeyan Shanmugam +1
We consider the problem of online regret minimization in linear bandits with access to prior observations (offline data) from the underlying bandit model. There are numerous applic…