1 paper · 1 filter
Ruixuan Miao, Xu Lu, Cong Tian +2
Unlike the standard Reinforcement Learning (RL) model, many real-world tasks are non-Markovian, whose rewards are predicated on state history rather than solely on the current stat…