2 citations · 2 across the 1 of their papers we have counts for
1 paper
David O'Callaghan, Patrick Mannion
When developing reinforcement learning agents, the standard approach is to train an agent to converge to a fixed policy that is as close to optimal as possible for a single fixed r…