1 paper · 1 filter
Marc Höftmann, Jan Robine, Stefan Harmeling
Can we learn policies in reinforcement learning without rewards? Can we learn a policy just by trying to reach a goal state? We answer these questions positively by proposing a mul…