5 citations · 7 across the 5 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2022
Nearly Minimax Algorithms for Linear Bandits with Shared Representation
Jiaqi Yang, Qi Lei, Jason D. Lee +1
We give novel algorithms for multi-task and lifelong linear bandits with shared representation. Specifically, we consider the setting where we play linear bandits with dimensio…
cs.LG2020★ 1 cited
Fully Gap-Dependent Bounds for Multinomial Logit Bandit
Jiaqi Yang
We study the multinomial logit (MNL) bandit problem, where at each time step, the seller offers an assortment of size at most from a pool of items, and the buyer purchases…
cs.LG2020
Linear Bandits with Limited Adaptivity and Learning Distributional Optimal Design
Yufei Ruan, Jiaqi Yang, Yuan Zhou
Motivated by practical needs such as large-scale learning, we study the impact of adaptivity constraints to linear contextual bandits, a central problem in online active learning.…