1 paper
Jordan Coblin, Han Wang, Martha White +1
A key obstacle to deploying reinforcement learning in real-world systems is hyperparameter selection, particularly when simulators are unavailable and online experimentation is cos…