1 paper
Fin Amin, Nathaniel Dennler, Andreea Bobu
Robots learn reward functions from user demonstrations, but these rewards often fail to generalize to new environments. This failure occurs because learned rewards latch onto spuri…