1 paper · 1 filter
Jingxin Zhan, Yuchen Xin, Kaicheng Jin +1
We study a stochastic convex bandit problem where the subgaussian noise parameter is assumed to decrease linearly as the learner selects actions closer and closer to the minimizer…