3 citations · 6 across the 6 of their papers we have counts for
3 papers · 1 filter
Omega-Regular Decision Processes
Ernst Moritz Hahn, Mateo Perez, Sven Schewe +3
Regular decision processes (RDPs) are a subclass of non-Markovian decision processes where the transition and reward functions are guarded by some regular property of the past (a l…
Reward Shaping for Reinforcement Learning with Omega-Regular Objectives
E. M. Hahn, M. Perez, S. Schewe +3
Recently, successful approaches have been made to exploit good-for-MDPs automata (Büchi automata with a restricted form of nondeterminism) for model free reinforcement learning, a…
Omega-Regular Objectives in Model-Free Reinforcement Learning
Ernst Moritz Hahn, Mateo Perez, Sven Schewe +3
We provide the first solution for model-free reinforcement learning of ω-regular objectives for Markov decision processes (MDPs). We present a constructive reduction from the almos…