2 citations · 2 across the 1 of their papers we have counts for
1 paper
Gabor Matuz, Andras Lorincz
Reinforcement learning has solid foundations, but becomes inefficient in partially observed (non-Markovian) environments. Thus, a learning agent -born with a representation and a p…