1 paper
Joel Q. L. Chang
We prove that ρ-NPTSSG, an anchor-free nonparametric Thompson Sampling algorithm for risk-averse bandits, achieves regret matching the instance-depend…