1 citations · 1 across the 3 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
ParMod: A Parallel and Modular Framework for Learning Non-Markovian Tasks
Ruixuan Miao, Xu Lu, Cong Tian +2
The commonly used Reinforcement Learning (RL) model, MDPs (Markov Decision Processes), has a basic premise that rewards depend on the current state and action only. However, many r…
cs.LG2023
Using Experience Classification for Training Non-Markovian Tasks
Ruixuan Miao, Xu Lu, Cong Tian +2
Unlike the standard Reinforcement Learning (RL) model, many real-world tasks are non-Markovian, whose rewards are predicated on state history rather than solely on the current stat…