1 paper · 1 filter
Pavel Osinenko, Grigory Yaremenko, Georgiy Malaniya +2
Reinforcement learning is commonly concerned with problems of maximizing accumulated rewards in Markov decision processes. Oftentimes, a certain goal state or a subset of the state…