1 paper
Meichen Song, Yuhao Wang, Enlu Zhou
In online reinforcement learning, data scarcity creates epistemic uncertainty that makes robustness important early in learning, whereas sufficient exploration is needed to learn t…