1 paper
Mahsa Bastankhah, Sophie Broderick, Benjamin Eysenbach
In many practical reinforcement learning environments, observations are far higher-dimensional than the variables that matter for control. In this work, we ask: can we learn repres…