31 citations · 31 across the 1 of their papers we have counts for
3 papers
Linear Bandits with Feature Feedback
Urvashi Oswal, Aniruddha Bhargava, Robert Nowak
This paper explores a new form of the linear bandit problem in which the algorithm receives the usual stochastic rewards as well as stochastic feedback about which features are rel…
Scalable Generalized Linear Bandits: Online Computation and Hashing
Kwang-Sung Jun, Aniruddha Bhargava, Robert Nowak +1
Generalized Linear Bandits (GLBs), a natural extension of the stochastic linear bandits, has been popular and successful in recent years. However, existing GLBs scale poorly with t…
Active Algorithms For Preference Learning Problems with Multiple Populations
Aniruddha Bhargava, Ravi Ganti, Robert Nowak
In this paper we model the problem of learning preferences of a population as an active learning problem. We propose an algorithm can adaptively choose pairs of items to show to us…