1 paper
Ori Linial, Guy Tennenholtz, Uri Shalit
In many reinforcement learning (RL) applications one cannot easily let the agent act in the world; this is true for autonomous vehicles, healthcare applications, and even some reco…