114 citations · 114 across the 1 of their papers we have counts for
1 paper · 1 filter
Philip J. Ball, Cong Lu, Jack Parker-Holder +1
Reinforcement learning from large-scale offline datasets provides us with the ability to learn policies without potentially unsafe or impractical exploration. Significant progress…