1 paper
Marco Lehmann, He Xu, Vasiliki Liakoni +3
In many daily tasks we make multiple decisions before reaching a goal. In order to learn such sequences of decisions, a mechanism to link earlier actions to later reward is necessa…