1 paper · 1 filter
Leonid Peshkin, Christian R. Shelton
Searching the space of policies directly for the optimal policy has been one popular method for solving partially observable reinforcement learning problems. Typically, with each c…