48 citations · 89 across the 7 of their papers we have counts for
1 paper · 1 filter
Taewoon Kim, Vincent François-Lavet, Michael Cochez
Partially observable reinforcement learning requires deciding what to retain, retrieve, and forget over time. We introduce a neuro-symbolic meta-policy that learns which symbolic m…