1 paper
Carlos E. Luis, Alessandro G. Bottero, Julia Vinogradska +2
Optimal decision-making under partial observability requires reasoning about the uncertainty of the environment's hidden state. However, most reinforcement learning architectures h…