3 citations · 4 across the 2 of their papers we have counts for
3 papers
Optimal Best-Arm Identification in Bandits with Access to Offline Data
Shubhada Agrawal, Sandeep Juneja, Karthikeyan Shanmugam +1
Learning paradigms based purely on offline data as well as those based solely on sequential online learning have been well-studied in the literature. In this paper, we consider com…
Regret Minimization in Heavy-Tailed Bandits
Shubhada Agrawal, Sandeep Juneja, Wouter M. Koolen
We revisit the classic regret-minimization problem in the stochastic multi-armed bandit setting when the arm-distributions are allowed to be heavy-tailed. Regret minimization has b…
City-Scale Agent-Based Simulators for the Study of Non-Pharmaceutical Interventions in the Context of the COVID-19 Epidemic
Shubhada Agrawal, Siddharth Bhandari, Anirban Bhattacharjee +14
We highlight the usefulness of city-scale agent-based simulators in studying various non-pharmaceutical interventions to manage an evolving pandemic. We ground our studies in the c…