8 citations · 8 across the 3 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Replicable Bandits with UCB based Exploration
Rohan Deb, Udaya Ghai, Karan Singh +1
We study replicable algorithms for stochastic multi-armed bandits (MAB) and linear bandits with UCB (Upper Confidence Bound) based exploration. A bandit algorithm is -replicabl…
cs.LG2026★ 8 cited
Introduction to Online Control
Elad Hazan, Karan Singh
This text presents an introduction to an emerging paradigm in control of dynamical systems and differentiable reinforcement learning called online nonstochastic control. The new ap…
cs.LG2025
Sample-Optimal Agnostic Boosting with Unlabeled Data
Udaya Ghai, Karan Singh
Boosting provides a practical and provably effective framework for constructing accurate learning algorithms from inaccurate rules of thumb. It extends the promise of sample-effici…