1 paper · 1 filter
Long-Fei Li, Peng Zhao, Zhi-Hua Zhou
We study reinforcement learning with linear function approximation, unknown transition, and adversarial losses in the bandit feedback setting. Specifically, we focus on linear mixt…