8 citations · 14 across the 2 of their papers we have counts for
4 papers
Explicit Best Arm Identification in Linear Bandits Using No-Regret Learners
Mohammadi Zaki, Avi Mohan, Aditya Gopalan
We study the problem of best arm identification in linearly parameterised multi-armed bandits. Given a set of feature vectors a confidence paramet…
Throughput Optimal Decentralized Scheduling with Single-bit State Feedback for a Class of Queueing Systems
Avinash Mohan, Aditya Gopalan, Anurag Kumar
Motivated by medium access control for resource-challenged wireless Internet of Things (IoT), we consider the problem of queue scheduling with reduced queue state information. In p…
On the Volatility of Optimal Control Policies and the Capacity of a Class of Linear Quadratic Regulators
Avinash Mohan, Shie Mannor, Arman Kizilkale
It is well known that highly volatile control laws, while theoretically optimal for certain systems, are undesirable from an engineering perspective, being generally deleterious to…
Towards Optimal and Efficient Best Arm Identification in Linear Bandits
Mohammadi Zaki, Avinash Mohan, Aditya Gopalan
We give a new algorithm for best arm identification in linearly parameterised bandits in the fixed confidence setting. The algorithm generalises the well-known LUCB algorithm of Ka…