1 paper
Quentin Delfosse, Sebastian Sztwiertnia, Mark Rothermel +2
Goal misalignment, reward sparsity and difficult credit assignment are only a few of the many issues that make it difficult for deep reinforcement learning (RL) agents to learn opt…