4 citations · 7 across the 12 of their papers we have counts for
1 paper · 1 filter
Pavel Osinenko, Grigory Yaremenko, Georgiy Malaniya +2
Reinforcement learning is commonly concerned with problems of maximizing accumulated rewards in Markov decision processes. Oftentimes, a certain goal state or a subset of the state…