1 paper · 1 filter
Hao Zhu, Jasper Hoffmann, Baohe Zhang +1
We consider the problem of fitting a reinforcement learning (RL) model to some given behavioral data under a multi-armed bandit environment. These models have received much attenti…