1 paper
Marcin Szulc, Jakub Łyskawa, Paweł Wawrzyński
The subject of this paper is reinforcement learning. Policies are considered here that produce actions based on states and random elements autocorrelated in subsequent time instant…