1 paper · 1 filter
Jean Tarbouriech, Evrard Garcelon, Michal Valko +2
Many popular reinforcement learning problems (e.g., navigation in a maze, some Atari games, mountain car) are instances of the episodic setting under its stochastic shortest path (…