5 citations · 6 across the 11 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023★ 1 cited
Stochastic Contextual Bandits with Long Horizon Rewards
Yuzhen Qin, Yingcong Li, Fabio Pasqualetti +2
The growing interest in complex decision-making and language modeling problems highlights the importance of sample-efficient learning over very long horizons. This work takes a ste…
cs.LG2022
Non-Stationary Representation Learning in Sequential Linear Bandits
Yuzhen Qin, Tommaso Menara, Samet Oymak +2
In this paper, we study representation learning for multi-task decision-making in non-stationary environments. We consider the framework of sequential linear bandits, where the age…