8 citations · 8 across the 1 of their papers we have counts for
1 paper
Alex Beeson, Giovanni Montana
The ability to discover optimal behaviour from fixed data sets has the potential to transfer the successes of reinforcement learning (RL) to domains where data collection is acutel…