1 paper
Eli Friedman, Fred Fontaine
Many reinforcement-learning researchers treat the reward function as a part of the environment, meaning that the agent can only know the reward of a state if it encounters that sta…