1 paper · 1 filter
Rasool Fakoor, Murdock Aubry, Nicholas Stranges +1
Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fragility is especially visibl…