6 citations · 9 across the 9 of their papers we have counts for
Showing 2022Show all
3 papers · 1 filter
stat.ML2022★ 1 cited
Double Doubly Robust Thompson Sampling for Generalized Linear Contextual Bandits
Wonyoung Kim, Kyungbok Lee, Myunghee Cho Paik
We propose a novel contextual bandit algorithm for generalized linear rewards with an regret over rounds where is the minimum eigenvalue of th…
stat.ML2022
Squeeze All: Novel Estimator and Self-Normalized Bound for Linear Contextual Bandits
Wonyoung Kim, Myunghee Cho Paik, Min-hwan Oh
We propose a linear contextual bandit algorithm with regret bound, where is the dimension of contexts and isthe time horizon. Our proposed algorithm is…
stat.ML2022
Semi-Parametric Contextual Bandits with Graph-Laplacian Regularization
Young-Geun Choi, Gi-Soo Kim, Seunghoon Paik +1
Non-stationarity is ubiquitous in human behavior and addressing it in the contextual bandits is challenging. Several works have addressed the problem by investigating semi-parametr…