1.1k citations · 1.1k across the 11 of their papers we have counts for
1 paper · 1 filter
Veronica Chelu, Tom Zahavy, Arthur Guez +2
We work towards a unifying paradigm for accelerating policy optimization methods in reinforcement learning (RL) by integrating foresight in the policy improvement step via optimist…