14 citations · 14 across the 1 of their papers we have counts for
1 paper
Rohin Shah, Noah Gundotra, Pieter Abbeel +1
Our goal is for agents to optimize the right reward function, despite how difficult it is for us to specify what that is. Inverse Reinforcement Learning (IRL) enables us to infer r…