1 paper
Zihan Ye, Phil Chau, Raban Emunds +5
Deep reinforcement learning agents are often misaligned, as they over-exploit early reward signals. Recently, several symbolic approaches have addressed these challenges by encodin…