1 paper
Gavin Rens, Jean-François Raskin, Raphaël Reynouad +1
There are situations in which an agent should receive rewards only after having accomplished a series of previous tasks, that is, rewards are non-Markovian. One natural and quite g…