15 citations · 15 across the 1 of their papers we have counts for
1 paper
Yecheng Jason Ma, Dinesh Jayaraman, Osbert Bastani
Many reinforcement learning (RL) problems in practice are offline, learning purely from observational data. A key challenge is how to ensure the learned policy is safe, which requi…