1 paper · 1 filter
Bohao Qu, Xiaofeng Cao, Jielong Yang +4
Markov Decision Process (MDP) presents a mathematical framework to formulate the learning processes of agents in reinforcement learning. MDP is limited by the Markovian assumption…