1 paper
Mohamad H. Danesh, Maxime Wabartha, Stanley Wu +2
Deploying reinforcement learning (RL) policies in real-world involves significant challenges, including distribution shifts, safety concerns, and the impracticality of direct inter…