1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Nevan Wichers
In the real world, RL agents should be rewarded for fulfilling human preferences. We show that RL agents implicitly learn the preferences of humans in their environment. Training a…