1 paper · 1 filter
Isaac J. Sledge, Matthew S. Emigh, Jose C. Principe
Reinforcement learning in environments with many action-state pairs is challenging. At issue is the number of episodes needed to thoroughly search the policy space. Most convention…