12 citations · 12 across the 1 of their papers we have counts for
1 paper
John D. Co-Reyes, Suvansh Sanjeev, Glen Berseth +2
Much of the current work on reinforcement learning studies episodic settings, where the agent is reset between trials to an initial state distribution, often with well-shaped rewar…