22 citations · 39 across the 6 of their papers we have counts for
Showing 2019Show all
2 papers · 1 filter
stat.ML2019
No-Regret Exploration in Goal-Oriented Reinforcement Learning
Jean Tarbouriech, Evrard Garcelon, Michal Valko +2
Many popular reinforcement learning problems (e.g., navigation in a maze, some Atari games, mountain car) are instances of the episodic setting under its stochastic shortest path (…
stat.ML2019★ 22 cited
Active Exploration in Markov Decision Processes
Jean Tarbouriech, Alessandro Lazaric
We introduce the active exploration problem in Markov decision processes (MDPs). Each state of the MDP is characterized by a random value and the learner should gather samples to e…