32 citations · 43 across the 4 of their papers we have counts for
Showing 2019Show all
2 papers · 1 filter
cs.LG2019★ 7 cited
Autonomous exploration for navigating in non-stationary CMPs
Pratik Gajane, Ronald Ortner, Peter Auer +1
We consider a setting in which the objective is to learn to navigate in a controlled Markov process (CMP) where transition probabilities may abruptly change. For this setting, we p…
cs.LG2019
Variational Regret Bounds for Reinforcement Learning
Pratik Gajane, Ronald Ortner, Peter Auer
We consider undiscounted reinforcement learning in Markov decision processes (MDPs) where both the reward functions and the state-transition probabilities may vary (gradually or ab…