3 citations · 7 across the 6 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2021★ 3 cited
Mungojerrie: Reinforcement Learning of Linear-Time Objectives
Ernst Moritz Hahn, Mateo Perez, Sven Schewe +3
Reinforcement learning synthesizes controllers without prior knowledge of the system. At each timestep, a reward is given. The controllers optimize the discounted sum of these rewa…
cs.LG2021
Model-free Reinforcement Learning for Branching Markov Decision Processes
Ernst Moritz Hahn, Mateo Perez, Sven Schewe +3
We study reinforcement learning for the optimal control of Branching Markov Decision Processes (BMDPs), a natural extension of (multitype) Branching Markov Chains (BMCs). The state…