1 citations · 1 across the 1 of their papers we have counts for
1 paper
Wendelin Böhmer, Rong Guo, Klaus Obermayer
This paper investigates a type of instability that is linked to the greedy policy improvement in approximated reinforcement learning. We show empirically that non-deterministic pol…