1 paper
Jorge de Heuvel, Sebastian Müller, Marlene Wessels +3
End-to-end robot policies achieve high performance through neural networks trained via reinforcement learning (RL). Yet, their black box nature and abstract reasoning pose challeng…