23 citations · 23 across the 1 of their papers we have counts for
1 paper
Maxime Gasse, Damien Grasset, Guillaume Gaudron +1
Learning efficiently a causal model of the environment is a key challenge of model-based RL agents operating in POMDPs. We consider here a scenario where the learning agent has the…